https://openreview.net/forum?id=aLLuYpn83y
Inference-Time Intervention: Eliciting Truthful Answers from a Language Model | OpenReview
We introduce Inference-Time Intervention (ITI), a technique designed to enhance the "truthfulness" of large language models (LLMs). ITI operates by shifting...
inference timelanguage modelinterventionelicitingtruthful
https://steerable-scene-generation.github.io/
Steerable Scene Generation with Post Training and Inference-Time Search
Steerable Scene Generation with Post Training and Inference-Time Search
post traininginference timescenegenerationsearch
https://arxiv.org/html/2508.21016v1
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance
inference timealignment controldiffusion modelsreinforcement learning
https://openreview.net/forum?id=fQUQPcTk8c
Enhancing Retrieval Systems with Inference-Time Logical Reasoning | OpenReview
Traditional retrieval methods rely on transforming user queries into vector representations and retrieving documents based on cosine similarity within an...
retrieval systemsinference timelogical reasoningenhancingopenreview
https://openreview.net/forum?id=EkQRNLPFcn
Towards Inference-time Category-wise Safety Steering for Large Language Models | OpenReview
While large language models (LLMs) have seen unprecedented advancements in capabilities and applications across a variety of use-cases, safety alignment of...
large language modelsinference timecategory wise
https://openreview.net/forum?id=1Uem0nAWK0&referrer=%5Bthe%20profile%20of%20MONICA%20SUNKARA%5D(%2Fprofile%3Fid%3D~MONICA_SUNKARA1)
Inference time LLM alignment in single and multidomain preference spectrum | OpenReview
Aligning Large Language Models (LLM) to address subjectivity and nuanced preference levels requires adequate flexibility and control, which can be a...
inference timellmalignment
https://www.amazon.science/blog/using-teacher-knowledge-at-inference-time-to-enhance-student-model
Using teacher knowledge at inference time to enhance student model - Amazon Science
May 28, 2024 - New method improves the state of the art in knowledge distillation by leveraging a knowledge base of teacher predictions.
inference time
https://arxiv.org/html/2511.07362v1
Inference-Time Scaling of Diffusion Models for Infrared Data Generation
inference timediffusion modelsscalinginfrareddata
https://openreview.net/forum?id=5wuZyG1ACs&referrer=%5Bthe%20profile%20of%20Etash%20Kumar%20Guha%5D(%2Fprofile%3Fid%3D~Etash_Kumar_Guha1)
Archon: An Architecture Search Framework for Inference-Time Techniques | OpenReview
Inference-time techniques are emerging as highly effective tools to enhance large language model (LLM) capabilities. However, best practices for developing...
architecture searchinference timearchonframeworktechniques
https://www.amazon.science/publications/detecting-hallucinations-in-speechllms-at-inference-time-using-attention-maps
Detecting hallucinations in SpeechLLMs at inference time using attention maps - Amazon Science
Hallucinations in Speech Large Language Models (SpeechLLMs) pose significant risks, yet existing detection methods typically rely on goldstandard outputs that...
detecting hallucinationsinference time
https://arxiv.org/abs/2409.15254
[2409.15254] Archon: An Architecture Search Framework for Inference-Time Techniques
Abstract page for arXiv paper 2409.15254: Archon: An Architecture Search Framework for Inference-Time Techniques
architecture searchinference time2409archon
https://arxiv.org/abs/2602.05285
[2602.05285] Robust Inference-Time Steering of Protein Diffusion Models via Embedding Optimization
Abstract page for arXiv paper 2602.05285: Robust Inference-Time Steering of Protein Diffusion Models via Embedding Optimization
https://arxiv.org/abs/2602.16138v2
[2602.16138v2] IRIS: Intent Resolution via Inference-time Saccades for Open-Ended VQA in Large...
Abstract page for arXiv paper 2602.16138v2: IRIS: Intent Resolution via Inference-time Saccades for Open-Ended VQA in Large Vision-Language Models
https://www.thoughtworks.com/en-ec/insights/blog/data-engineering/accelerating-real-time-inference
Accelerating real-time inference | Thoughtworks Ecuador
It is crucial that real-time and batch inference share the same data stores to minimize defects during the feature extraction process. While batch and...
real timeacceleratinginferencethoughtworksecuador