Intelligence DEV
The Dirty Secret Behind AI Agents (Demo 🚀)
For quite a while now, I've had the feeling that AI agents are surrounded by this mystical aura....
Full corpus search
25 matches for “agents”
Intelligence DEV
For quite a while now, I've had the feeling that AI agents are surrounded by this mystical aura....
Intelligence DEV
Run two AI coding agents on the same repo and the first thing that breaks is not the code. It is...
Intelligence arXiv
Modern AI agents rely on elaborate inference harnesses such as Claude Code, Codex, and OpenClaw to drive multi-turn reasoning, tool use, and access to external systems. While powerful, these complex harnesses also make agents hard to train end-to-end with…
Intelligence DEV
Hi friends Here's a tip for when you're setting up Claude Code (or any coding agent) in a real...
Intelligence Hacker News
awsmux fans one AWS CLI command out across hundreds of accounts and regions in parallel and merges the results into a single stream. There's an MCP server built in for agents.In a 150-session benchmark, agents using awsmux beat agents using the raw AWS CLI in…
Intelligence arXiv
Real-time multimodal applications, including voice agents and interactive video generation, compose heterogeneous models into pipelines whose efficient deployment requires application-specific decisions about placement, streaming, and intra-model parallelism…
Intelligence arXiv
Coding agents increasingly operate in executable environments where a failed attempt produces actionable feedback rather than merely an incorrect answer. Existing cost-aware systems typically treat such failures as cascade decisions: try a cheap model first…
Intelligence arXiv
Agentic systems large language model (LLM) based architectures capable of reasoning, planning, acting, and coordinating with tools and other agents are rapidly transitioning from research prototypes to production scale deployments across domains such as…
Intelligence DEV
In Agentic interaction using AppFunctions I showed how Be nice publishes createAppPair for agents,...
Intelligence arXiv
As AI agents begin to automate AI R&D, we need ways to assess whether their outputs are safe to deploy, even when the agents themselves may be untrusted. AI control offers one such approach: rather than trusting the agent, it treats it as a potential…
Intelligence Hacker News
https://x.com/jack/status/2079605800998146171, https://xcancel.com/jack/status/2079605800998146171https://buzz.xyz/
Intelligence Ars Technica
Augment Code's Vinay Perneti talks models, harnesses, and context.
Intelligence Ars Technica
Aggressive training techniques sharpens threat of bad behavior by leading models.
Intelligence DEV
There is a distinct, quiet disquiet settling over the software engineering landscape today. You can...
Engineering DEV
We have plenty of posts on DEV announcing new tools, libraries, agents, and frameworks. I want to...
Intelligence DEV
"You're absolutely right!" The agent says it before it has looked at anything. You point out a bug,...
Intelligence Ars Technica
"This is day one for cybersecurity in the age of agents," Hugging Face CEO says.
Intelligence DEV
AI may help junior developers ship faster. It may also slow down how they become senior. That...