ZeroHour

Search: “OpenAL”

6 stories in the last 24h

[AINews] not much happened today

Latent Space AI news digest covers Anthropic's Claude Code Projects, Google's managed agent APIs, TypeSafe's Jev classifier, and OpenAI's Astra for Law launch.

The 9/16-9/17/2026 AI news roundup highlights Anthropic's Claude Code Projects enabling one conversation to spawn parallel cloud sessions, and Google's Gemini managed agents adding a Credentials API, Files API, and claims of 30% lower costs. It also covers TypeSafe's Jev, a fast constrained-output classifier being used for routing, judgment, and structured decisions, with open reproductions such as openjev-s on Qwen3.6-35B-A3B. OpenAI launched Astra for Law with 26 partner-built and 47 community plugins via Trusted Access, with reports it beats generic GPT-6 Astra plus web search on Vals' legal benchmark. Research items include DeepMind's Stellar Colosseum multi-agent math harness (Codeforces 4263, 71.0% on TCS-Bench) and NVIDIA-associated Agora using Git commits as shared memory.

Anthropic keeps pushing Claude Code toward autonomous coding with new parallel agent workflows

Anthropic rebuilt Claude Code's Projects feature to split goals across parallel cloud agent threads that can open pull requests and run tests.

Anthropic's updated Claude Code Projects feature uses a coordinator that splits a user's goal into parallel threads, each running as its own cloud session, with shared memory and a library of uploaded files and results. Threads can open pull requests and run tests, and progress is trackable in the main chat or per thread, including on mobile. The beta is open to select Pro and Max subscribers, with Team and Enterprise access and local execution to follow. The update follows Anthropic making autopilot mode the default in Claude Code.

The Decoder · 19h agoAI industry

How to Write with an LLM

Thomas Ptacek publishes a method for LLM-assisted writing: never adopt suggested words and forbid model encouragement to preserve author voice.

Thomas Ptacek outlines two rules for using LLMs as copyeditors: never use a single word a model suggests, and forbid encouragement that reinforces first-draft impulses. He recommends running model passes to flag passive voice, repetition, and misplaced paragraphs, and comparing rewrites with a fresh-context model to avoid bias. He also recommends the book 'Style: Lessons in Clarity and Grace' and mentions building a small tool to manage context-free copyediting comparisons.

Hacker News · AIupdated · 14h agofirst · 16h agoAI industry 2 sourcesHN 46↑ · 31 comments

Salesforce Agentforce: Bridging the Enterprise AI Gap from ‘Vibe Coding’ to Battle-Tested Orchestration

Salesforce pitches Agentforce as an enterprise agent platform with testing, observability, and deterministic gating; Southwest Airlines reports $6M annual savings and 45% autonomous resolution.

Salesforce positions Agentforce as an enterprise agent harness built on Data Cloud and Customer 360, exposing external endpoints via the Model Context Protocol and offering Agentforce Testing Center for synthetic stress-testing, headless CI/CD regressions, Agent Optimizer for live prompt tuning, and deterministic gating to prevent unvalidated actions like payments. Southwest Airlines deployed Agentforce across its Help Center and mobile app starting November 2025, reporting a 45% autonomous resolution rate across more than 2 million interactions, 7x ROI, $6 million in projected annual savings, and a +900% jump in customer satisfaction metrics. The article frames the platform as competing with other enterprise agent orchestration offerings.

MarkTechPost · 6h agoAI industry1

UN turns to Google to make its global data ready for AI agents

UN and Google launch UN System Data Commons for AI agents; UNICEF benchmark finds six major LLMs answered global statistics questions with just 21.2% accuracy.

The UN launched the UN System Data Commons, built on Google's open-source Data Commons, replacing the UNData portal and supporting MCP so AI agents can query authoritative statistics with source traceability. A UNICEF benchmark of over 133,000 responses found GPT-4o, GPT-4o-mini, Claude Sonnet 4.5, Claude Haiku 4.5, Gemini 2.5 Flash, and Gemini 2.0 Flash averaged just 21.2% accuracy on global development indicators, with roughly three in five answers providing no usable number. Google.org contributed $2 million in funding, 26 UN entities have committed to the platform, and ChatGPT referrals to UNICEF's data site rose 67% year over year.

TechCrunch · AI · 18h agoAI industry 2 sources

Pinterest teases a new ‘Restyle’ feature that lets you redesign your room with AI

Pinterest launched Restyle, a beta AI feature in the US and Canada that redesigns rooms from photos via Pinterest Intelligence.

Pinterest introduced Restyle in beta in the US and Canada, letting users visualize furniture, decor, paint colors, and style changes in photos of their own rooms. The feature is powered by Pinterest Intelligence, built on Nvidia Blackwell GPUs and Nvidia Dynamo combined with open-source models and Pinterest-built technology. It was teased at the Pinterest Presents event alongside visual search ads, with broader rollout planned next month.

TechCrunch · AI · 20h agoAI industry