Robuta

https://openreview.net/forum?id=aLLuYpn83y Inference-Time Intervention: Eliciting Truthful Answers from a Language Model | OpenReview We introduce Inference-Time Intervention (ITI), a technique designed to enhance the "truthfulness" of large language models (LLMs). ITI operates by shifting... inference timelanguage modelinterventionelicitingtruthful https://steerable-scene-generation.github.io/ Steerable Scene Generation with Post Training and Inference-Time Search Steerable Scene Generation with Post Training and Inference-Time Search post traininginference timescenegenerationsearch https://arxiv.org/html/2508.21016v1 Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance inference timealignment controldiffusion modelsreinforcement learning https://openreview.net/forum?id=fQUQPcTk8c Enhancing Retrieval Systems with Inference-Time Logical Reasoning | OpenReview Traditional retrieval methods rely on transforming user queries into vector representations and retrieving documents based on cosine similarity within an... retrieval systemsinference timelogical reasoningenhancingopenreview https://openreview.net/forum?id=EkQRNLPFcn Towards Inference-time Category-wise Safety Steering for Large Language Models | OpenReview While large language models (LLMs) have seen unprecedented advancements in capabilities and applications across a variety of use-cases, safety alignment of... large language modelsinference timecategory wise https://openreview.net/forum?id=1Uem0nAWK0&referrer=%5Bthe%20profile%20of%20MONICA%20SUNKARA%5D(%2Fprofile%3Fid%3D~MONICA_SUNKARA1) Inference time LLM alignment in single and multidomain preference spectrum | OpenReview Aligning Large Language Models (LLM) to address subjectivity and nuanced preference levels requires adequate flexibility and control, which can be a... inference timellmalignment https://www.amazon.science/blog/using-teacher-knowledge-at-inference-time-to-enhance-student-model Using teacher knowledge at inference time to enhance student model - Amazon Science May 28, 2024 - New method improves the state of the art in knowledge distillation by leveraging a knowledge base of teacher predictions. inference time https://arxiv.org/html/2511.07362v1 Inference-Time Scaling of Diffusion Models for Infrared Data Generation inference timediffusion modelsscalinginfrareddata https://openreview.net/forum?id=5wuZyG1ACs&referrer=%5Bthe%20profile%20of%20Etash%20Kumar%20Guha%5D(%2Fprofile%3Fid%3D~Etash_Kumar_Guha1) Archon: An Architecture Search Framework for Inference-Time Techniques | OpenReview Inference-time techniques are emerging as highly effective tools to enhance large language model (LLM) capabilities. However, best practices for developing... architecture searchinference timearchonframeworktechniques https://www.amazon.science/publications/detecting-hallucinations-in-speechllms-at-inference-time-using-attention-maps Detecting hallucinations in SpeechLLMs at inference time using attention maps - Amazon Science Hallucinations in Speech Large Language Models (SpeechLLMs) pose significant risks, yet existing detection methods typically rely on goldstandard outputs that... detecting hallucinationsinference time https://arxiv.org/abs/2409.15254 [2409.15254] Archon: An Architecture Search Framework for Inference-Time Techniques Abstract page for arXiv paper 2409.15254: Archon: An Architecture Search Framework for Inference-Time Techniques architecture searchinference time2409archon https://arxiv.org/abs/2602.05285 [2602.05285] Robust Inference-Time Steering of Protein Diffusion Models via Embedding Optimization Abstract page for arXiv paper 2602.05285: Robust Inference-Time Steering of Protein Diffusion Models via Embedding Optimization https://arxiv.org/abs/2602.16138v2 [2602.16138v2] IRIS: Intent Resolution via Inference-time Saccades for Open-Ended VQA in Large... Abstract page for arXiv paper 2602.16138v2: IRIS: Intent Resolution via Inference-time Saccades for Open-Ended VQA in Large Vision-Language Models https://www.thoughtworks.com/en-ec/insights/blog/data-engineering/accelerating-real-time-inference Accelerating real-time inference | Thoughtworks Ecuador It is crucial that real-time and batch inference share the same data stores to minimize defects during the feature extraction process. While batch and... real timeacceleratinginferencethoughtworksecuador