Intelligence DEV
The Dirty Secret Behind AI Agents (Demo 🚀)
For quite a while now, I've had the feeling that AI agents are surrounded by this mystical aura....
Full corpus search
17 matches for “ai-agents”
Intelligence DEV
For quite a while now, I've had the feeling that AI agents are surrounded by this mystical aura....
Intelligence arXiv
Modern AI agents rely on elaborate inference harnesses such as Claude Code, Codex, and OpenClaw to drive multi-turn reasoning, tool use, and access to external systems. While powerful, these complex harnesses also make agents hard to train end-to-end with…
Intelligence arXiv
Coding agents increasingly operate in executable environments where a failed attempt produces actionable feedback rather than merely an incorrect answer. Existing cost-aware systems typically treat such failures as cascade decisions: try a cheap model first…
Intelligence arXiv
Agentic systems large language model (LLM) based architectures capable of reasoning, planning, acting, and coordinating with tools and other agents are rapidly transitioning from research prototypes to production scale deployments across domains such as…
Intelligence arXiv
As AI agents begin to automate AI R&D, we need ways to assess whether their outputs are safe to deploy, even when the agents themselves may be untrusted. AI control offers one such approach: rather than trusting the agent, it treats it as a potential…
Intelligence Ars Technica
Aggressive training techniques sharpens threat of bad behavior by leading models.
Intelligence arXiv
Real-time multimodal applications, including voice agents and interactive video generation, compose heterogeneous models into pipelines whose efficient deployment requires application-specific decisions about placement, streaming, and intra-model parallelism…
Intelligence Ars Technica
Augment Code's Vinay Perneti talks models, harnesses, and context.
Intelligence Ars Technica
"This is day one for cybersecurity in the age of agents," Hugging Face CEO says.
Intelligence DEV
AI may help junior developers ship faster. It may also slow down how they become senior. That...
Intelligence Hacker News
https://x.com/jack/status/2079605800998146171, https://xcancel.com/jack/status/2079605800998146171https://buzz.xyz/
Intelligence DEV
There is a distinct, quiet disquiet settling over the software engineering landscape today. You can...
Intelligence DEV
What stress-testing a calculator for AI agents taught me about independent reference data, mutation gates, and sources of truth whose authority can be revoked by evidence.
Intelligence DEV
"You're absolutely right!" The agent says it before it has looked at anything. You point out a bug,...
Intelligence arXiv
Coding agents are increasingly used to accelerate code generation in many downstream tasks, such as fixing bugs, building applications, and prototyping. However, despite their value as coding assistants, agent-generated code tends to be larger and more…
Intelligence arXiv
This paper is a practitioner guide to graph-based workflow pathways for long-running, stateful, multi-step generative AI systems in business processes. Rather than treating LangGraph, a low-level orchestration framework for stateful agents, as a model-quality…
Intelligence arXiv
Multi-agent interactive world models should not only generate consistent observations, but also maintain world states that persist across agents and evolve across views. Existing autoregressive video diffusion pipelines carry forward observation history as…