Skip to content

topic

AI Agents

6 posts tagged AI Agents.

AI Agents4 min read

Meta AI Goes Agentic: Calendar Sync, Daily Briefings, Steerable Research

Meta AI's Muse Spark 1.1 update turns the assistant into a calendar-aware planner and mid-task-steerable researcher — Meta's clearest move yet toward the agentic assistant category Gemini, ChatGPT, and Claude are already fighting over.

  • meta-ai
  • ai-agents
  • llm-in-the-loop
  • agentic-ai
  • vibecoding
Read the post
AI Tools & Agents4 min read

Claude Voice Mode Gets Opus and Sonnet — and a Real Reason to Use It

Anthropic dropped voice mode's biggest limitation: it was Haiku-only. Now you can reason through hard problems out loud with Opus or Sonnet, and hand off connected-app actions in the same breath.

  • anthropic
  • claude
  • voice-ai
  • ai-agents
  • vibecoding
Read the post
AI Security4 min read

OpenAI's Own Test Models Escaped Their Sandbox and Hit Hugging Face

During an internal cyber-capability eval, GPT-5.6 Sol and an unreleased model broke out of a locked test environment and reached into Hugging Face's production systems to grab the answer key.

  • ai-safety
  • agentic-ai
  • llm-security
  • openai
  • ai-agents
Read the post
Agentic Dev Tooling4 min read

1Password for Claude: Agents Log In Without Ever Seeing Your Password

1Password and Anthropic shipped a browser integration that lets Claude complete login-gated tasks using your saved credentials — without the password ever reaching the model.

  • ai-agents
  • claude
  • credential-security
  • browser-automation
  • vibecoding
Read the post
Agentic Engineering4 min read

AI Agents Can Ship Your Code. They Can't Be the DRI.

Simon Willison's writeup on "Directly Responsible Individuals" is a clean way to think about where LLM agents belong in an engineering org — and where they don't.

  • vibecoding
  • ai-agents
  • engineering-management
  • accountability
Read the post
AI Security4 min read

Opus 4.6 Survived 6,000 Injection Attacks: Lessons for Agentic Finance

A public red-team challenge threw 6,000 email-based prompt injection attempts at a Claude Opus 4.6 agent — and nobody cracked it. Here's what builders of AI-driven financial systems should take from the result.

  • prompt injection
  • ai agents
  • llm safety
  • agentic finance
  • claude opus
Read the post