https://studylib.net/doc/27711720/2505.22617v1
RL Entropy Mechanism for Reasoning LLMs: Collapse & Control
This paper investigates policy entropy collapse in RL for LLMs, revealing a predictable performance-entropy trade-off. It analyzes entropy dynamics and...
reasoning llmsrlentropymechanismcollapse
https://twimlai.com/podcast/twimlai/ai-trends-2026-openclaw-agents-reasoning-llms
AI Trends 2026: OpenClaw Agents, Reasoning LLMs, and More | TWIML - The Voice of Machine Learning &...
Feb 26, 2026 - In this episode, Sebastian Raschka, independent LLM researcher and author, joins us to break down how the LLM landscape has changed over the past year and what...
ai trendsreasoning llmsand morethe voicemachine learning
https://www.gpttranslator.co/blog/from-neural-mt-to-reasoning-llms-gpt-translator-evolution
From Neural MT to Reasoning LLMs: The Next Evolution in GPT Translator
Through this blog, the paper explores its history from the Neural Machine Translation (NMT) phase to becoming a reasoning large language model. It drives the...
reasoning llmsthe nextgpt translatorneuralmt
https://huggingface.co/papers/2410.03645
Paper page - GenSim2: Scaling Robot Data Generation with Multi-modal and Reasoning LLMs
Join the discussion on this paper page
data generationreasoning llmspaperscalingrobot
https://scipapermill.com/2025/10/27/unleashing-ais-inner-thinker-recent-advances-in-chain-of-thought-reasoning-for-llms-and-beyond/
Unleashing AI's Inner Thinker: Recent Advances in Chain-of-Thought Reasoning for LLMs and Beyond
Dec 28, 2025 - Latest 50 papers on chain-of-thought reasoning: Oct. 27, 2025
chain of thoughtrecent advancesfor llmsand beyondinner
https://appleinsider.com/articles/24/10/12/apples-study-proves-that-llm-based-ai-models-are-flawed-because-they-cannot-reason
Reasoning failures highlighted by Apple research on LLMs
A new paper from Apple's artificial intelligence scientists has found that engines based on large language models, such as those from Meta and OpenAI, still...
by applereasoningfailureshighlightedresearch
https://www.datacamp.com/ko/tutorial/chain-of-thought-prompting
Chain-of-Thought Prompting: Step-by-Step Reasoning with LLMs | DataCamp
Unlock the full potential of Large Language Models (LLMs) with our guide on Chain-of-Thought (CoT) prompting.
chain of thoughtpromptingstepreasoningllms
https://tldr.takara.ai/p/2504.19095
Efficient Reasoning for LLMs through Speculative Chain-of-Thought | Takara TLDR
Large reasoning language models such as OpenAI-o1 and Deepseek-R1 have recently attracted widespread attention due to their impressive task-solving abilities...
chain of thoughtfor llmsefficientreasoningspeculative
https://www.aibase.com/tool/31091
Visual Sketchpad-A visual reasoning tool for multimodal large language models (LLMs)
Visual Sketchpad is a framework that provides a visual sketchpad and drawing tools for multimodal large language models (LLMs). It allows models to operate on v
large language modelsvisualsketchpadreasoningtool
https://data.mendeley.com/datasets/97wy9y3t6y/2
Istidlal: A Dataset for Benchmarking Logical and Sequential Economic Reasoning in Arabic LLMs -...
This repository contains the data and resources associated with Istidlal, the first Arabic benchmark for evaluating logical and sequential reasoning in Arabic...
logical anddatasetbenchmarkingsequentialeconomic
https://arxiv.org/abs/2506.03295
[2506.03295] Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One...
Abstract page for arXiv paper 2506.03295: Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem
fine tuningon onereasoningpotentialpre
https://openreview.net/forum?id=9ljHiYuRHl&referrer=%5Bthe%20profile%20of%20Gaurav%20Rohit%20Ghosal%5D(%2Fprofile%3Fid%3D~Gaurav_Rohit_Ghosal1)
Failure Modes of LLMs for Causal Reasoning on Narratives | OpenReview
In this work, we investigate the causal reasoning abilities of large language models (LLMs) through the representative problem of inferring causal...
failure modesllmscausalreasoningnarratives
https://papers.neurips.cc/paper_files/paper/2025/hash/1b6a42d0966442d77e1ceee22bfad49b-Abstract-Conference.html
Matching Markets Meet LLMs: Algorithmic Reasoning with Ranked Preferences
matching marketsmeetllmsalgorithmicreasoning
https://flexa.careers/in/jobs/sap-phd-student-agentic-ai-knowledge-graphs-llms-autonomous-query-reasoning-f-m-69f8e8ca2257ee3f1b6b9bf8
PhD Student : Agentic AI : Knowledge Graphs, LLMs & Autonomous Query Reasoning F/M at SAP
Join SAP Labs France as a PhD Student in Agentic AI. Work on knowledge graphs, LLMs, and autonomous reasoning in a 3-year research-focused role.
phd studentagentic aiknowledge graphsf mllms
https://sol.sbc.org.br/index.php/webmedia_estendido/article/view/38209/37984
Vista do Evaluating Zero-shot Reasoning with Agentic LLMs for Smart Contract Vulnerability Detection
zero shotagentic llmssmart contractvulnerability detectionvista
https://chatpaper.com/paper/276129
Strat-Reasoner: Reinforcing Strategic Reasoning of LLMs in Multi-Agent Games
Strat-Reasoner is a novel reinforcement learning framework that enhances the strategic reasoning capabilities of Large Language Models in multi-agent games by...
multi agentstratreinforcingreasoningllms
https://liner.com/review/beyond-answers-transferring-reasoning-capabilities-to-smaller-llms-using-multiteacher
Beyond Answers: Transferring Reasoning Capabilities to Smaller LLMs Using Multi-Teacher Knowledge...
Regarding this WSDM 2024 paper, this review summarizes TINYLLM, a novel knowledge distillation paradigm for transferring diverse reasoning capabilities to ...
beyondanswerstransferringreasoningcapabilities
https://www.catalyzex.com/paper/limited-reasoning-space-the-cage-of-long
Limited Reasoning Space: The cage of long-horizon reasoning in LLMs
Limited Reasoning Space: The cage of long-horizon reasoning in LLMs: Paper and Code. The test-time compute strategy, such as Chain-of-Thought (CoT), has...
the cagelimitedreasoningspacelong