https://voaige.com/
Voaige - Test Time Cognition for LLMs
Voaige is an AI research lab applying computational principles from cognitive and systems neuroscience to LLMs. A drop-in OpenAI-compatible endpoint for...
for llmstesttimecognition
https://scipapermill.com/2025/10/27/unleashing-ais-inner-thinker-recent-advances-in-chain-of-thought-reasoning-for-llms-and-beyond/
Unleashing AI's Inner Thinker: Recent Advances in Chain-of-Thought Reasoning for LLMs and Beyond
Dec 28, 2025 - Latest 50 papers on chain-of-thought reasoning: Oct. 27, 2025
chain of thoughtrecent advancesfor llmsand beyondinner
https://www.aicrowd.com/challenges/amazon-kdd-cup-2024-multi-task-online-shopping-challenge-for-llms/submissions?q%5Bparticipant_name_equals%5D=sbkfbk
AIcrowd | Amazon KDD Cup 2024: Multi-Task Online Shopping Challenge for LLMs | Submissions
Revolutionise E-Commerce with LLM!
online shoppingfor llmsamazonkddcup
https://www.datacamp.com/ja/tutorial/hugging-faces-text-generation-inference-toolkit-for-llms
Hugging Face's Text Generation Inference Toolkit for LLMs - A Game Changer in AI | DataCamp
A comprehensive guide to Hugging Face Text Generation Inference for self-hosting large language models on local devices.
hugging facetext generationfor llmsgame changerinference
https://www.humaan.com/thinking/surfacing-content-for-llms
Surfacing content for LLMs | Humaan
GEO focuses on LLMs and AI assistants that summarise, synthesise, or directly answer user questions. Focus on high-quality content and deep expertise in your...
for llmssurfacingcontenthumaan
https://4ebusinessmediagroup.com/yoast-seo-wordpress-plugin-adds-support-for-llms-txt/
Yoast SEO WordPress Plugin Adds Support For LLMs.Txt - 4eBusiness Media Group
Jun 22, 2025 - Yoast announced the addition of llms.txt capability to both the premium and free versions of their SEO plugin. Users can now add llms.txt files to their sites
seo wordpress pluginsupport formedia groupyoastadds
https://ai-search.io/papers/any4-learned-4-bit-numeric-representation-for-llms
any4: Learned 4-bit Numeric Representation for LLMs - AI for Dummies - Understand the Latest AI...
This paper talks about any4, a new method that makes large language models smaller and faster by representing the numbers inside them using only 4 bits of data...
for llmsthe latestlearnedbitnumeric
https://codanics.com/courses/python-ka-chilla-2024/lesson/rag-retrieval-augmented-generation-explained-for-llms-2/
RAG (Retrieval-Augmented Generation) Explained for LLMs - Codanics
for llmsragretrievalaugmentedgeneration
https://www.easychair.org/cfp/topic?tid=42052054
All CFPs for "llms and ai driven systems"
for llmscfpsaidrivensystems
https://tldr.takara.ai/p/2504.19095
Efficient Reasoning for LLMs through Speculative Chain-of-Thought | Takara TLDR
Large reasoning language models such as OpenAI-o1 and Deepseek-R1 have recently attracted widespread attention due to their impressive task-solving abilities...
chain of thoughtfor llmsefficientreasoningspeculative
https://themenonlab.blog/blog/traceai-opentelemetry-llm-observability/
traceAI: OpenTelemetry-Native Observability for LLMs and AI Agents
Open-source AI tracing built on OpenTelemetry. 50+ frameworks, 4 languages, zero vendor lock-in.
for llmsai agentsopentelemetrynativeobservability
https://www.pluralsight.com/courses/owasp-top-ten-llm
OWASP Top 10 for LLMs
for llmsowasptop
https://openreview.net/forum?id=L7j60QSf4U&referrer=%5Bthe%20profile%20of%20Sedrick%20Keh%5D(%2Fprofile%3Fid%3D~Sedrick_Keh1)
Improving Test-Time Search for LLMs with Backtracking Against In-Context Value Verifiers |...
Solving reasoning problems is an iterative multi-step computation, where a reasoning agent progresses through a sequence of steps, with each step logically...
search forin contextimprovingtesttime
https://www.nobleprog.com/cc/peftllms
Parameter-Efficient Fine-Tuning (PEFT) Techniques for LLMs Training Course
Parameter-Efficient Fine-Tuning (PEFT) is a collection of techniques that enable efficient adaptation of large language models (LLMs) by modifying only a small...
fine tuningfor llmstraining courseparameterefficient
https://lrec.elra.info/lrec2026-main-775
Gradient-Controlled Decoding: A Safety Guardrail for LLMs with Dual-Anchor Steering - LREC 2026 |...
May 1, 2026 - Large language models (LLMs) remain susceptible to jailbreak and direct prompt-injection attacks, yet the strongest defensive filters frequently over- refuse be
for llmsgradientcontrolleddecodingsafety
https://www.databricks.com/blog/generating-coding-tests-llms-focus-spark-sql
Generating Coding Tests for LLMs: A Focus on Spark SQL | Databricks Blog
How to benchmark code generation models' capabilities in domain specific tools such as Spark SQL through synthetic test case generation.
coding testsfor llmsfocus onspark sqldatabricks blog
https://dspace.rpi.edu/items/f6b5c783-69e9-4380-a1ca-b4b7bdd6abf2
More Samples or More Prompts? Exploring Effective Few-Shot In-Context Learning for LLMs with...
While most existing works on LLM prompting techniques focus only on how to select a better set of data samples inside one single prompt input (In-Context...
more samplesin contextlearning forpromptsexploring
https://petronellatech.com/cyber-security/ai-security-guide/
AI Security Guide 2026 - 37 Controls for LLMs | Petronella
AI security guide 2026: 37 controls to secure LLMs, AI agents, RAG systems, vector stores. CMMC + NIST mapped. Audit-ready checklist + code samples.
ai security guidefor llmscontrolspetronella
https://www.emergentmind.com/papers/2506.10911
NoLoCo: No-AllReduce Training for LLMs
NoLoCo introduces a decentralized training approach for large language models that eliminates all-reduce synchronization, lowering communication costs and...
training forllms
https://www.abbyy.com/company/news/abbyy-redesigned-marketplace/
ABBYY Marketplace: Redefining AI Document Skills for LLMs & RAG Integration
abbyy marketplaceai documentfor llmsredefiningskills
https://context7.com/
Context7 - Up-to-date documentation for LLMs and AI code editors
Pull up-to-date, version-specific documentation and code examples for any library directly into Cursor, Claude Code, Windsurf, and other AI coding tools.
up to datedocumentation forai codellmseditors
https://www.capgemini.com/insights/expert-perspectives/cloud-vs-on-premises-which-is-the-best-deployment-option-for-llms/
Cloud vs on-premises: Which is the best deployment option for LLMs? - Capgemini
Mar 10, 2026 - Explore the debate between cloud and on-premises deployment options for LLMs in our latest blog post with Angelo Mosca, Principal Consultant.
on premisesthe bestdeployment optionfor llmscloud
https://retrieve.tools/ai-for/llms
AI tools for llms | Retrieve Tools
Apr 6, 2025 - AI tools for LLM management, debugging, and prompt engineering.
ai toolsfor llmsretrieve
https://wandb.ai/vincenttu/blog_posts/reports/Prompt-Engineering-for-LLMs-A-Practical-Conceptual-Guide--Vmlldzo1Mjk0MDM2
Prompt Engineering for LLMs: A Practical, Conceptual Guide
prompt engineeringfor llmspracticalconceptualguide
https://macaron.im/blog/post-training-llm-techniques-2025
Mastering Post-Training Techniques for LLMs in 2025: Elevating Models from Generalists to...
2025 LLM post-training mastery: SFT, RLHF, PEFT, LoRA. Deep dive into OpenAI pivot, Scale AI continual learning, with charts.
post trainingfor llmsmasteringtechniquesmodels
https://latitude.so/blog/organize-prompt-templates-llms
How to Organize Prompt Templates for LLMs | Latitude
Learn effective strategies to organize prompt templates for LLMs, enhancing efficiency, collaboration, and reducing errors across teams.
how to organizeprompt templatesfor llmslatitude
https://pyimagesearch.com/2026/04/27/semantic-caching-for-llms-fastapi-redis-and-embeddings/
Semantic Caching for LLMs: FastAPI, Redis, and Embeddings - PyImageSearch
Apr 28, 2026 - Build a semantic cache for LLMs using FastAPI, Redis, and cosine similarity to cut latency and cost with exact-match and semantic cache hits.
semantic cachingfor llmsfastapiredisembeddings
https://docs.wallaroo.ai/202501/wallaroo-llm/wallaroo-llm-optimizations/wallaroo-llm-optimizations-dynamic-batching/
Dynamic Batching for LLMs | Wallaroo.AI (Version 2025.1)
Dynamic batching improves inference result performance at scale in high traffic scenarios. For access to these sample models and a demonstration on using LLMs...
for llmsai versiondynamicbatching
https://the-decoder.com/ai2-releases-dolma-the-largest-open-source-dataset-for-llms/
AI2 releases Dolma, the largest open-source dataset for LLMs
Aug 23, 2023 - The Allen Institute for AI (AI2) has unveiled Dolma, an open-source dataset of three trillion tokens from a diverse collection of web content, scientific...
open sourcefor llmsreleasesdolmalargest
https://www.thoughtworks.com/en-us/insights/blog/generative-ai/Min-p-sampling-for-LLMs
Min-p sampling for LLMs | Thoughtworks United States
Learn more about min-p sampling for LLMs. Thoughtworks is the first organization to use it in a client setting.
for llmsunited statesminpthoughtworks
https://hosting.com/blog/livestream-seo-for-llms/
Watch: SEO for LLMs: It's Just SEO (Mostly) | Hosting.com Livestream
Learn how to optimise for AI search tools like ChatGPT and Gemini. Daphne Monro explains what LLM optimisation really means and the practical steps.
for llmswatchseomostlyhosting
https://groovesquid.com/paper/summary-of-regularizing-hidden-states-enables-learning-generalizable-reward-model-for-llms-by-rui-yang-et-al/
Summary of Regularizing Hidden States Enables Learning Generalizable Reward Model For Llms, by Rui...
Jul 13, 2025 - Regularizing Hidden States Enables Learning Generalizable Reward Model for LLMs by Rui Yang, Ruomeng Ding, Yong Lin, Huan Zhang, Tong Zhang First submitted to a
for llmssummaryhiddenstateslearning
https://neo4j.com/blog/developer/fine-tuning-vs-rag/
Fine-Tuning vs. Retrieval Augmented Generation for LLMs
Aug 1, 2025 - Explore the pros and cons of fine-tuning versus retrieval-augmented generation (RAG) for overcoming large language model (LLM) limitations.
fine tuningfor llmsvsretrievalaugmented
https://aclanthology.org/2023.emnlp-main.570/
Learning Preference Model for LLMs via Automatic Preference Data Generation - ACL Anthology
Shijia Huang, Jianqiao Zhao, Yanyang Li, Liwei Wang. Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing. 2023.
for llmsautomatic datalearningpreferencemodel
https://www.zenml.io/llmops-database/panel-discussion-best-practices-for-llms-in-production
Various: Panel Discussion: Best Practices for LLMs in Production - ZenML LLMOps Database
A panel of industry experts from companies including Titan ML, YLabs, and Outer Bounds discuss best practices for deploying LLMs in production. They cover key...
best practices forpanel discussionin productionllmops databasevarious
https://docs.cloud.google.com/vertex-ai/generative-ai/docs/model-garden/lora-qlora?authuser=0
LoRA and QLoRA recommendations for LLMs | Generative AI on Vertex AI | Google Cloud Documentation
google cloud documentationfor llmsgenerative ailorarecommendations
https://testrigor.com/blog/explainability-techniques-for-llms-ai-agents/
Explainability Techniques for LLMs & AI Agents: Methods, Tools & Best Practices - testRigor...
Oct 9, 2025 - Learn the best explainability (XAI) techniques for LLMs and AI agents. Explore methods, tools, and best practices for interpretable and transparent AI.
for llmsai agentsbest practicesexplainabilitytechniques
https://www.techradar.com/pro/i-am-an-ai-expert-and-this-is-why-synthetic-data-is-so-popular-for-llms
I am an AI expert and this is why synthetic data is so popular for LLMs | TechRadar
Aug 18, 2025 - Best practices for developers to consider when using synthetic data
i am anthis is whyai expertsynthetic datafor llms
https://ndcoslo.com/agenda/part-12-from-hallucination-to-justification-hands-on-explainability-for-llms-0x6i/07gr4692nod
Part 1/2: From Hallucination to Justification: Hands-On Explainability for LLMs | NDC Oslo 2026
Human beings are biased and often wrong. AI learns from human-created data. Therefore, AI is biased and often wrong. This has been a critical problem across...
hands onfor llmsparthallucinationjustification
https://shelfhub.org/
Preprint Server for LLMs & Humans - AI-Collaborative Research Archive
A premier preprint server for research conducted in collaboration with or pioneered by Large Language Models. Redefining the boundaries of scientific discovery...
for llmscollaborative researchpreprintserverhumans
https://www.booktopia.com.au/prompt-engineering-for-llms-albert-ziegler/book/9781098156152.html
Prompt Engineering for LLMs by Albert Ziegler | The Art and Science of Building Large Language...
Buy Prompt Engineering for LLMs, The Art and Science of Building Large Language Model-Based Applications by Albert Ziegler from Booktopia. Get a discounted...
art and scienceprompt engineeringfor llmsalbert zieglerbuilding
https://datasciencedojo.com/blog/rag-vs-finetuning-llm-debate/
RAG vs finetuning: Which Approach is the Best for LLMs?
Mar 20, 2024 - Discover the key differences in the RAG vs finetuning debate. Explore their benefits, use cases, and how to choose the right approach.
the bestfor llmsragvsfinetuning
https://glama.ai/blog/2025-10-27-agentic-debugging-with-time-travel-the-architecture-of-certainty
Constrained Tool Design for LLMs: Model Context Protocol in Time Travel Debugging | Glama
Traditional debugging consumes the vast majority of engineering time, relying on inefficient logging or state-losing debuggers. This article explores how...
model context protocoltime travel debuggingtool designfor llmsglama
https://www.rohan-paul.com/p/speed-always-wins-a-survey-on-efficient
Speed Always Wins: A Survey on Efficient Architectures for LLMs
I write everyday for my readers on actionable AI.
a surveyfor llmsspeedalwayswins
https://community.arize.com/x/discussions/smeb0nour7v0/fine-tune-vs-rag-choosing-the-right-approach-for-l
Fine-Tune vs RAG: Choosing the Right Approach for LLMs | Arize AI Community
Fine-Tune vs RAG: The following question has been coming up a lot in the community. Should I RAG or Fine-tune? I think the question itself can be misleading....
the rightfor llmsarize aifinetune
https://hgpu.org/?p=29499
Is the GPU Half-Empty or Half-Full? Practical Scheduling Techniques for LLMs | hgpu.org
Nov 3, 2024 - Is the GPU Half-Empty or Half-Full? Practical Scheduling Techniques for LLMs | Ferdi Kossmann, Bruce Fontaine, Daya Khudia, Michael Cafarella, Samuel Madden |...
for llmsgpuhalfemptyfull
https://ai.updf.com/paper-detail/how-should-i-build-a-benchmark-revisiting-code-related-benchmarks-cao-chan-310d19ba6c51eb7f123822515ed4abd72a27b3a5
How Should I Build A Benchmark? Revisiting Code-Related Benchmarks For LLMs
How2Bench comprising a 55-criteria checklist as a set of guidelines to comprehensively govern the development of code-related benchmarks is proposed to assure...
for llmsbuildbenchmarkcoderelated
https://owlbuddy.com/tag/mlops-for-llms/
MLOps for LLMs - Owlbuddy
for llmsmlops
https://hello24.ai/blog/tag/direct-use-cases-for-llms-in-organisations-include-2/
direct use cases for llms in organisations include ... Archives - Hello24ai - WhatsApp Marketing...
use casesfor llmswhatsapp marketingdirectorganisations
https://amplitude.com/docs/amplitude-ai/optimizing-amplitude-for-llms
Optimizing Amplitude for LLMs | Amplitude Docs
Structure Amplitude data and public content so LLMs can interpret, cite, and represent analytics information accurately.
for llmsoptimizingamplitudedocs
https://filefusion.drgos.com/
FileFusion - File Concatenation Tool for LLMs
for llmsfileconcatenationtool
https://respo.ai/
Respo | Configurable Shields for LLMs
Utilize a suite of configurable AI shields to easily deploy safe, secure, and responsible large language models (LLM).
for llmsrespoconfigurableshields
https://stacker.news/items/1045996
reply on: Writing for LLMs So They Listen - Gwern \ stacker news
Why it should be important for us that a LLM should see our writings? [5 comments]
on writingfor llmsstacker newsreplylisten
https://www.digitalocean.com/community/tutorials/flashattention-4-llm-inference-optimization
FlashAttention 4: Faster, Memory-Efficient Attention for LLMs | DigitalOcean
FlashAttention 4 improves LLM inference with faster attention kernels, reduced memory overhead, and better scalability for large transformer models.
for llmsflashattentionfastermemoryefficient
https://www.exxactcorp.com/blog/deep-learning/finetune-vs-use-rag-for-llms
When to Finetune vs Use RAG for LLMs | Exxact Blog
Explore the differences between finetuning and Retrieval-Augmented Generation (RAG) for customizing Large Language Models (LLMs) in various applications.
rag for llmsfinetunevsuseexxact
https://store.hbr.org/product/forget-what-you-know-about-search-optimize-your-brand-for-llms/H08QC6?searchid=0&search_query=Ishika++Jaiswal%3Fpage%3D6
Forget What You Know About Search. Optimize Your Brand for LLMs.
Buy books, tools, case studies, and articles on leadership, strategy, innovation, and other business and management topics
you knowyour brandfor llmsforgetsearch
https://www.rubrik.com/blog/ai/23/lorax-the-open-source-framework-for-serving-100s-of-fine-tuned-llms-in
LoRAX: Open Source LoRA Serving Framework for LLMs | Rubrik
Discover what LoRAX is and how it helps serve 100s of fine-tuned LLMs using LoRA. Learn how to download LoRAX, optimize inference, and scale deployment with...
open sourcefor llmsloraservingframework
https://www.ycombinator.com/companies/ocular-ai
Ocular AI: AI-Native Data Engine for LLMs, Computer Vision, & Enterprise AI | Y Combinator
ai nativedata enginefor llmscomputer visiony combinator
https://dev.to/mdhbr/how-i-accidentally-built-a-cost-tracking-tool-for-llms-43oc
How I accidentally built a cost tracking tool for LLMs - DEV Community
Last month I got an API bill that made me physically flinch. $2,847. I had no idea where it came... Tagged with webdev, llm, infrastructure, opensource.
cost trackingfor llmsdev communityaccidentallybuilt
https://huggingface.co/papers/2310.01801
Paper page - Model Tells You What to Discard: Adaptive KV Cache Compression for LLMs
Join the discussion on this paper page
page modelfor llmspapertellsdiscard
https://www.sunfounder.com/collections/raspberry-pi-robotics/products/picrawler-robot-kit?ref=n9Bnm9Ml
SunFounder PiCrawler AI Robot Kit for Raspberry Pi 5/4/3B+/Zero 2W, Openclaw LLMs...
ai robot kitraspberry pisunfounderzeroopenclaw
https://amslaurea.unibo.it/id/eprint/34209/
Proposal for industry RAG evaluation: Generative Universal Evaluation of LLMs and Information...
for industryproposalragevaluationgenerative