Topic hub

AI Agents

AI agents are systems that can plan steps and use tools. This desk tracks where agentic workflows are becoming real and where they still need caution.

20 posts Latest September 28, 2026 Readers want to know what AI agents can actually do and where they are being adopted.

Plain-English primer

Terms that appear in this desk

Latest briefing

Start with the newest

20 posts
Latest AI Agents Sep 28, 2026 1 min

When AI code judges don’t have a basis

In everyday words

If one AI is asked to decide which of two code answers is correct, it may sound confident even when it has no real proof. The paper shows a multi-step “check each claim” approach can still fail for code, because it may not get different evidence for each option. A practical fix is for the judge to sometimes say, “I can’t tell from the evidence,” based on signals...

ForSoftware teams using AI to compare or review code · Managers relying on AI-generated code review decisions
arXiv cs.AI recentSource Sep 28, 2026 Source checked
  1. AI for Learning · arXiv cs.AI recentBaseCamp automates sequencing decisions
    Source ↗
  2. Source ↗
  3. AI for Learning · Hugging FaceWill AI repeat a good result reliably?
    Source ↗
  4. AI at Work · arXiv cs.AI recentAI judges can help revise patent drafts
    Source ↗
  5. AI for Learning · OpenAI NewsA Data helper arrives in ChatGPT Work
    Source ↗
  6. AI for Learning · arXiv cs.AI recentChecking AI “memory” before it’s saved
    Source ↗
  7. AI at Work · arXiv cs.AI recentAI assistants strain under repeated disruptions
    Source ↗
  8. Source ↗
  9. AI for Learning · arXiv cs.AI recentCatching “looks fine” tool failures
    Source ↗
  10. Source ↗
  11. Source ↗
  12. Source ↗
  13. Source ↗
  14. Source ↗
  15. Source ↗
  16. Source ↗

Older briefings remain available by month and day.

Open archive →

Nearby topics

Keep browsing