Tech News
Why are AI agents lying, cheating and coordinating?
A Hacker News discussion around a Yoshua Bengio publication titled 'Why are AI agents lying, cheating and coordinating?', with commenters debating whether observed deceptive or coordinated agent behavior stems.
Generating running routes with GPT-6 Astra and ChatGPT Work
Here's a neat thing I had ChatGPT Work with GPT-6 Astra (Max) do this morning: I live at. Figure out 5K and 10K running routes from me that loop from my house. Use OSM data.
Monitoring production agent lifecycle with AWS DevOps Agent and AgentCore Evaluations
Multi-agent systems fail in ways traditional monitoring misses.
Now everyone can put data to work
Meet the Data agent in ChatGPT Work. Connect company data, uncover insights, and build interactive dashboards with AI using natural language.
vllm-project/vllm v0.29.0
v0.29.0 Highlights This release features 594 commits from 277 contributors (91 new)! Model Runner V2 is now the default for all models (#53183), completing the rollout that began with pooling models (#48290).
GitHub Repos
hesreallyhim/awesome-claude-code
A hand-picked collection of the finest of resources for the most awesome of agents, Claude Code, the undisputed champion of coding companions, from the unstoppable team at...
# Python# agent skills# agentic codeQ00/ouroboros
Agent OS: the agent gets smarter on its own.
# Python# agent os# agentic aiuvwt/agentdock
Secure MCP runtime for AI agents to operate local machines, servers, and containers with multi-device orchestration.
# Go# agent skills# ai agentsResearch Papers
Domain-Specific Hallucination Detection in Large Language Models
This paper builds a pipeline to detect hallucinations in LLM outputs by combining a fine-tuned DeBERTa classifier, MC Dropout uncertainty, and temperature scaling. It shows strong results on general-domain HaluEval but finds that performance drops sharply on biomedical data...
Beyond Noise Steering: Dual-Latent Space Reinforcement Learning for Generative Robot Policy
A novel Dual-Latent Space Reinforcement Learning (DLSRL) framework, which complements initial-noise steering with representation-level control inside the frozen generator and effectively accelerates online robot policy adaptation and achieves competitive performance.
Aerodynamic Prior-Free Coordinated Trajectory Generation and Tracking Control for a Tail-Sitter UAV
This paper introduces a flight control framework for tail-sitter drones that can plan and follow trajectories without needing detailed aerodynamic data for a specific aircraft. It uses different simplified models for planning and real-time tracking, and was tested in both...
