https://huggingface.co/papers/2402.06664
Paper page - LLM Agents can Autonomously Hack Websites
Join the discussion on this paper page
paper pagellm agentsautonomouslyhackwebsites
https://aisc.substack.com/p/llm-agents-part-6-state-management/comments
Comments - LLM Agents, Part 6 - State Management
How can we control the behavior of agents by drawing lines around the boundaries of their agency?
llm agentscommentspartstatemanagement
https://arxiv.org/abs/2502.04358
[2502.04358] Position: Scaling LLM Agents Requires Asymptotic Analysis with LLM Primitives
Abstract page for arXiv paper 2502.04358: Position: Scaling LLM Agents Requires Asymptotic Analysis with LLM Primitives
llm agentspositionscaling
https://arxiv.org/abs/2404.08144v2
[2404.08144v2] LLM Agents can Autonomously Exploit One-day Vulnerabilities
Abstract page for arXiv paper 2404.08144v2: LLM Agents can Autonomously Exploit One-day Vulnerabilities
llm agentsone dayautonomouslyexploitvulnerabilities
https://arxiv.org/abs/2507.21504
[2507.21504] Evaluation and Benchmarking of LLM Agents: A Survey
Abstract page for arXiv paper 2507.21504: Evaluation and Benchmarking of LLM Agents: A Survey
llm agentsevaluationbenchmarkingsurvey
https://arxiv.org/abs/2603.08640v1
[2603.08640v1] PostTrainBench: Can LLM Agents Automate LLM Post-Training?
Abstract page for arXiv paper 2603.08640v1: PostTrainBench: Can LLM Agents Automate LLM Post-Training?
llm agentsautomateposttraining
https://par.nsf.gov/biblio/10637439-evaluating-llm-agents-simulating-humanoid-behavior
Evaluating the LLM Agents for Simulating Humanoid Behavior | NSF Public Access Repository
This page contains metadata information for the record with PAR ID 10637439
llm agents
https://docs.google.com/forms/d/e/1FAIpQLSevYR6VaYK5FkilTKwwlsnzsn8yI_rRLLqDZj0NH7ZL_sCs_g/closedform
LLM Agents MOOC Hackathon Participant Signup
Please join LLM Agents MOOC discord for more discussions about the hackathon, including finding potential teammates if you are interested. Please see LLM...
llm agentsmoochackathonparticipantsignup
https://fcrc.rpi.edu/research/ai-algorithms/systematic-failure-analysis-llm-agents-taxonomy-attribution-and-reflection
Systematic Failure Analysis for LLM Agents: Taxonomy, Attribution, and Reflection | Future of...
failure analysisfor llm
https://www.salesforce.com/uk/agentforce/llm-agents/?bc=OTH
LLM Agents: A Complete Guide | Salesforce UK
LLM agents can parse complicated questions, improve decision-making and take timely action. Here's a look at the types of LLM agents and their benefits.
a complete guidellm agentssalesforceuk
https://kyunnilee.github.io/projects/1_project/
A SCAVENGER HUNT GAME FOR LLM AGENTS | Heekyung (Anne) Lee
A benchmark to evaluate spatial reasoning and navigational capabilities of LLM agents via a scavenger hunt game. (Final project for CS194/294, UC Berkeley)
scavenger huntfor llmgameagentsanne
https://arxiv.org/abs/2507.05257v3
[2507.05257v3] Evaluating Memory in LLM Agents via Incremental Multi-Turn Interactions
Abstract page for arXiv paper 2507.05257v3: Evaluating Memory in LLM Agents via Incremental Multi-Turn Interactions
llm agents
https://deepsense.ai/case-studies/exploring-llm-agents-for-innovation-with-tailored-llm-workshops/
Exploring LLM Agents for Innovation with Tailored LLM Workshops - deepsense.ai
llm agentsfor innovationtailored workshopsexploring
https://www.anthropic.com/research/shade-arena-sabotage-monitoring
SHADE-Arena: Evaluating Sabotage and Monitoring in LLM Agents \ Anthropic
A new set of evaluations to test the sabotage and monitoring capabilities of LLM AI models
llm agentsshadearenaevaluatingsabotage
https://deepsense.ai/tech-expertise/llms-rag/
Custom LLM & RAG Solutions - AI Agents, Assistants, Copilots - ds.ai
Jul 29, 2026 - Unlock the power of LLMs and RAG with expert AI solutions. We develop custom AI agents, assistants, copilots, and enterprise AI applications to enhance...
custom llmrag solutionsai agentsassistantscopilots
https://arize.com/blog/arize-nvidia-nemo-integration/
Self-Improving Agents: Automating LLM Performance Optimization using Arize and NVIDIA NeMo - Arize...
Jul 31, 2026 - The Arize integration of NVIDIA NeMo empowers AI teams with an automated, self-improving AI data flywheel to enhance LLM performance.
llm performance
https://techcommunity.microsoft.com/blog/azure-ai-foundry-blog/build-your-dream-team-with-autogen/4157961
Build a powerful team of LLM agents, to solve complicated multi step tasks
Jun 13, 2024 - LLM, Multi agents, Autogen, Azure Open AI
https://artinoid.com/
AI Development Company | LLM Applications, AI Agents, Agentic AI & Automation
AI development company in India building LLM applications, AI agents, and intelligent automation systems. We create scalable AI products for startups and...
ai development companyllm applicationsagentsagenticautomation
https://www.businesswire.com/news/home/20251030640798/en/Elastic-Brings-LLM-Observability-to-Azure-AI-Foundry-to-Optimize-AI-Agents
Elastic Brings LLM Observability to Azure AI Foundry to Optimize AI Agents
Elastic (NYSE: ESTC), the Search AI Company, today announced a new integration with Azure AI Foundry, delivering observability for agentic AI applications an...
azure ai foundryllm observabilityelasticbringsoptimize
https://botlab.dev/
botlab.dev | Custom AI/LLM Bots, Agents, Swarms, Automation
custom aibotlabdevllmbots
https://speakerdeck.com/phoenixhawk/basta-2025-agents-in-action-llms-tools-and-reasoning
BASTA! 2025: Agents in Action: LLM's, Tools and Reasoning - Speaker Deck
Slides for my talk about how to implement agents at BASTA! 2025 in Mainz.
in action
https://www.akamai.com/blog/developers/the-prompt-as-a-rulebook-guiding-llm-agents-beyond-basic-instructions
The Prompt as a Rulebook - Guiding LLM Agents Beyond Basic Instructions | Akamai
In our journey of integrating Large Language Models (LLMs) with traditional APIs, we've seen how prompts become the new "API docs," describing what a tool In...
the promptas a
https://arxiv.org/abs/2510.00615
[2510.00615] ACON: Optimizing Context Compression for Long-horizon LLM Agents
Abstract page for arXiv paper 2510.00615: ACON: Optimizing Context Compression for Long-horizon LLM Agents
context compressionaconoptimizing
https://adambossy.com/
Adam Bossy-Mendoza | Architecting LLM apps, Building Agents, 0 to 1 and 1 to N Startup Engineering...
A Medium-style blog built with Astro
https://matrices.app/
Matrices - Training Environments for LLM Agents
Training environments for multimodal LLM-based agents on realistic computer use tasks.
training environmentsfor llmmatricesagents
https://www.businesswire.com/news/home/20250717118204/en/Zoho-Launches-Zia-LLM-and-Deepens-AI-Portfolio-with-Prebuilt-Agents-Custom-Agent-Builder-MCP-and-Marketplace?
Zoho Launches Zia LLM and Deepens AI Portfolio with Prebuilt Agents, Custom Agent Builder, MCP, and...
Zoho Corporation, a global technology company, today announced additional investments and offerings in AI, including Zia LLM, a proprietary large language mo...
https://arxiv.org/html/2509.25885v1
SafeMind: Benchmarking and Mitigating Safety Risks in Embodied LLM Agents
safety riskssafemindbenchmarkingembodiedllm
https://gpbib.cs.ucl.ac.uk/gp-html/brookes2025evolvingexcellenceautomatedoptimization.html
Evolving Excellence: Automated Optimization of LLM-based Agents
automated optimizationevolvingexcellencellmbased
https://manifest.build/
The Open Source LLM Router for AI Agents | Manifest
The open source LLM router that connects your AI agents and harnesses to any provider in seconds. Subscriptions, pay-per-token, local models, and custom...
for ai agentsthe openllm routersourcemanifest
https://arxiv.org/abs/2601.20334?ref=swarmsignal.net
[2601.20334] Demonstration-Free Robotic Control via LLM Agents
Abstract page for arXiv paper 2601.20334: Demonstration-Free Robotic Control via LLM Agents
demonstrationfreeroboticcontrolvia
https://infoscience.epfl.ch/entities/publication/8e6e74ae-8ff6-48f8-918a-07cf9d92e4c1
PharmaSimText: A Text-Based Educational Playground filled with RL-LLM Agents That Work Together...
There has been a growing interest in developing simulated learners to enhance learning and teaching experiences in educational environments. However, existing...
https://dev.to/liuhaotian2024prog/why-auditing-ai-agents-requires-causal-ai-not-another-llm-269c
Why auditing AI agents requires causal AI, not another LLM - DEV Community
Your Logs Tell You What Happened. They Don't Tell You What Should Have Happened. Haotian... Tagged with ai, python, agents, opensource.
ai agentsauditingrequires
https://console.ngage360.app/
NGage 360 - Deploy LLM Powered AI Agents Across Devices
NGage 360 - Create custom LLM-powered agents to enhance customer engagement through personalized AI-driven interactions, turning every touchpoint into a sales...
ai agentsngagedeployllmpowered
https://ai.hku.hk/news-events/events/upcoming/eventdetail/507/36/rethinking-software-root-cause-analysis-with-llm-agents
Rethinking Software Root Cause Analysis with LLM Agents
root cause analysisrethinkingsoftwarellmagents
https://docs.recall.ai/docs/agent-quickstarts
Coding Agents/LLM Quickstarts
Jul 14, 2026 - Quick copy/pasteable instructions for Coding Agents/LLMs to implement transcription with bots. Coding Agents/LLMs are required to visit this page before...
coding agentsllmquickstarts
https://zenodo.org/records/21355030
Gubernaut: A Deterministic Homeostatic Controller for Affect-Regulated LLM Agents, Validated Across...
Jul 29, 2026 - The paper and its complete validation record for the Gubernaut Cognitive Controller (GCC), a model-agnostic, deterministic runtime control layer for LLM agents...
https://arxiv.org/abs/2505.14668v1
[2505.14668v1] ContextAgent: Context-Aware Proactive LLM Agents with Open-World Sensory Perceptions
Abstract page for arXiv paper 2505.14668v1: ContextAgent: Context-Aware Proactive LLM Agents with Open-World Sensory Perceptions
https://arxiv.org/abs/2510.24695v1
[2510.24695v1] AgentFrontier: Expanding the Capability Frontier of LLM Agents with ZPD-Guided Data...
Abstract page for arXiv paper 2510.24695v1: AgentFrontier: Expanding the Capability Frontier of LLM Agents with ZPD-Guided Data Synthesis
https://www.gonemo.ai/
Build AI Agents With Ease. nemo™ True No-code Multi-model LLM Platform
nemo™, a true no-code platform empowers midmarket teams with flexible AI solutions that integrate seamlessly into your workflow.
build ai agents
https://llmgraph.ai/home
No-code LLM workflow builder for RAG & AI agents - LLMGraph
Build LLM workflows as a graph, no code. Turn docs and models into RAG chatbots and AI agents, then ship to a REST API and embeddable chat widget. 14-day trial.
no codeworkflow builderrag aillmagents
https://scholars.cityu.edu.hk/en/publications/surveying-with-ai-simulating-human-responses-using-personalized-l/
Surveying with AI: Simulating Human Responses Using Personalized LLM Agents and Social Media Data -...
https://www.businesswire.com/news/home/20251028168283/en/Lakera-Launches-Open-Source-Security-Benchmark-for-LLM-Backends-in-AI-Agents
Lakera Launches Open-Source Security Benchmark for LLM Backends in AI Agents
Check Point Software Technologies Ltd. (NASDAQ: CHKP), a pioneer and global leader of cyber security solutions, and Lakera, a world leading AI-native securit...
open source security
https://openai.github.io/openai-agents-python/ref/extensions/models/any_llm_model/
Any-LLM model - OpenAI Agents SDK
openai agentsllmmodelsdk
https://huggingface.co/papers/2412.20138
Paper page - TradingAgents: Multi-Agents LLM Financial Trading Framework
Join the discussion on this paper page
paper pagemulti agentsfinancial tradingllmframework
https://arxiv.org/abs/2411.04671
[2411.04671] CUIfy the XR: An Open-Source Package to Embed LLM-powered Conversational Agents in XR
Abstract page for arXiv paper 2411.04671: CUIfy the XR: An Open-Source Package to Embed LLM-powered Conversational Agents in XR
https://researchdiscovery.drexel.edu/esploro/outputs/preprint/Orchestrating-LLM-Agents-for-Scientific-Research/991022164541904721?institution=01DRXU_INST&skipUsageReporting=true&recordUsage=false
Orchestrating LLM Agents for Scientific Research: A Pilot Study of Multiple Choice Question (MCQ)...
Advances in large language models (LLMs) are rapidly transforming scientific work, yet empirical evidence on how these systems reshape research activities...
https://ar5iv.labs.arxiv.org/html/2412.20138
[2412.20138] TradingAgents: Multi-Agents LLM Financial Trading Framework
Significant progress has been made in automated problem-solving using societies of agents powered by large language models (LLMs). In finance, efforts have...
multi agentsfinancial tradingllmframework
https://arxiv.org/abs/2601.00240
[2601.00240] When Agents See Humans as the Outgroup: Belief-Dependent Bias in LLM-Powered Agents
Abstract page for arXiv paper 2601.00240: When Agents See Humans as the Outgroup: Belief-Dependent Bias in LLM-Powered Agents
https://gubernaut.com/
Gubernaut · A deterministic cognitive governor for LLM agents
Gubernaut is a deterministic cognitive governor for LLM agents: a model-agnostic runtime AI control layer, measured across four frontier model families in a...
for llmdeterministiccognitivegovernoragents
https://arxiv.org/abs/2603.22341
[2603.22341] T-MAP: Red-Teaming LLM Agents with Trajectory-aware Evolutionary Search
Abstract page for arXiv paper 2603.22341: T-MAP: Red-Teaming LLM Agents with Trajectory-aware Evolutionary Search
https://bwrc.berkeley.edu/tokens-tapeout-llm-agents-tackling-hard-problems-modern-chip-design-may-8
From Tokens to Tapeout: LLM Agents Tackling the Hard Problems of Modern Chip Design | May 8 |...
https://www.udacity.com/course/langchain-agentic-ai-fundamentals--cd14639
Build LLM Apps & Agents with LangChain | Udacity
Gain hands-on skills creating LangChain chatbots and multi-step workflows. Explore agent design, memory, and state management to deploy intelligent AI agents.
buildllmappsagentslangchain
https://opensource.googleblog.com/2025/12/grl-turning-verifiable-games-into-a-post-training-suite-for-llm-agents-with-tunix-on-tpus.html?m=0
GRL: Turning verifiable games into a post-training suite for LLM agents with Tunix on TPUs | Google...
https://arxiv.org/abs/2407.03884
[2407.03884] ChatSOP: An SOP-Guided MCTS Planning Framework for Controllable LLM Dialogue Agents
Abstract page for arXiv paper 2407.03884: ChatSOP: An SOP-Guided MCTS Planning Framework for Controllable LLM Dialogue Agents