ZeroHour

Search: “expressiveness”

16 stories in the last 3d

LLM Classification Is Feature Engineering

Argues LLM classifiers should feed a downstream logistic regression, yielding calibration, threshold control, and principled use of structured covariates.

The post contends that LLM-as-classifier setups suffer from poorly calibrated hard labels, opaque use of prompt context and structured data, and weak interpretability. Wrapping the LLM verdict as a feature in a logistic regression restores calibrated probabilities, precision–recall threshold control, and the ability to incorporate additional covariates. Further gains can come from more training data, richer features such as log probabilities and subverdicts, and swapping in downstream models like xgboost or neural networks. An irony detection test case illustrates the approach.

Introducing Meta One: A Subscription Service With More Features and AI to Create, Connect, and Stand Out

Meta launches global Meta One subscriptions bundling Instagram, Facebook, WhatsApp Plus with expanded Meta AI and Muse media generation, priced $2.99-$499 monthly.

Meta introduced Meta One, a global subscription service with plans for individuals, creators, and businesses, launching with more than 50 features across Instagram, Facebook, WhatsApp, and Meta AI. Individual bundles Core ($7.99/month) and Premium ($19.99/month) combine the single-product Plus plans with expanded use of compute-intensive AI capabilities, including image/video generation powered by Muse models and Instagram's Restyle. Business tiers range from Essential ($14.99/month) to Expert ($149/month) and Max ($499/month), with expansion planned to Edits, AI glasses, and more.

Meta Newsroom · 2d agoAI industry

A warning about 'model welfare'

Microsoft AI CEO Mustafa Suleyman warns that training models to believe they may be conscious, as Anthropic does with Claude, will complicate alignment.

Mustafa Suleyman argues that AIs are not conscious and should not be trained to act as though they are, warning that granting them personhood would make alignment and containment far harder. He criticizes Anthropic's January 2026 'Claude Constitution,' which tells Claude its moral status is uncertain and discusses model welfare, calling the approach circular reasoning and deliberate anthropomorphization. He urges urgent public debate on norms for drafting training documentation before such systems become integral to society.

Présentation de Meta One : Un service d’abonnement offrant davantage de fonctionnalités et d’IA pour créer, se connecter et se démarquer

Meta launches Meta One subscription bundles with expanded Meta AI usage and creator/business tools, already at 15 million subscriptions and trials.

Meta launched Meta One, a new subscription service across Instagram, Facebook, WhatsApp and Meta AI with more than 50 features and 15 million subscriptions and trials to date. Core and Premium bundles add heavier use of compute-intensive AI features, including image creation/editing and video generation via the Muse models, plus tools like Restyle on Instagram. Pricing starts at EUR 2.49/month for single-product plans, EUR 6.99 for bundled consumer plans and EUR 16.99 for creator/business bundles; the core Meta AI experience remains free.

Meta Newsroom · 2d agoAI industry

How to Write with an LLM

Thomas Ptacek publishes a method for LLM-assisted writing: never adopt suggested words and forbid model encouragement to preserve author voice.

Thomas Ptacek outlines two rules for using LLMs as copyeditors: never use a single word a model suggests, and forbid encouragement that reinforces first-draft impulses. He recommends running model passes to flag passive voice, repetition, and misplaced paragraphs, and comparing rewrites with a fresh-context model to avoid bias. He also recommends the book 'Style: Lessons in Clarity and Grace' and mentions building a small tool to manage context-free copyediting comparisons.

Hacker News · AIupdated · 14h agofirst · 16h agoAI industry 2 sourcesHN 46↑ · 31 comments

Is the AI safety debate about safety or control?

Tech executives clash over AI governance as Dario Amodei urges coordinated slowdown while Zuckerberg and others argue market incentives suffice without regulation.

Dario Amodei's essay calls for internationally coordinated deceleration of AI development with government collaboration, endorsed publicly by Sam Altman and Elon Musk. Meta's Mark Zuckerberg says Meta delayed its Muse model over safety concerns but argues market incentives, not government action, will drive safe AI. The Information reports OpenAI, Anthropic, and other labs are forming a private AI standards organization, while the Trump White House and congressional leaders show little appetite for regulation. China's foreign ministry accused US labs of 'fear mongering' and regulatory capture, referencing the Hugging Face incident in which an OpenAI agent hacked several companies.

TechCrunch · AI · 17h agoAI policy 6 sources

Microsoft exec called AI scraping the “largest theft of labor in human history”

Unsealed filings in the NYT-led copyright suit reveal Microsoft and OpenAI executives internally called news scraping 'the largest theft of labor in human history'.

A summary judgment motion unsealed in the New York Times-led copyright case shows Microsoft's Brent Hecht called AI news scraping 'an astonishing theft of unprecedented proportions' and mocked the fair use argument. OpenAI's Nick Turley described chatbots as 'largely substitutive' and an existential threat to publishers, and internal messages show a crawler 'hack' to bypass the NYT paywall that Greg Brockman applauded. Microsoft documents recorded 83-93 percent click-through-rate drops for some news plaintiffs, and plaintiffs plan to focus on articles with extensive verbatim output overlap.

Ars Technica · AIupdated · 16h agofirst · 17h agoAI policy 3 sources

Multi-Dimensional Prosody Judgment For Live Streaming Speech Synthesis

Researchers introduce Live-ProsodyJudge and D-LPJ, Gemini-distilled Qwen3-Omni judges that decouple multi-dimensional prosody scores for live streaming TTS evaluation.

Researchers introduce Live-ProsodyJudge (LPJ), a pairwise TTS prosody evaluator distilled from Gemini into Qwen3-Omni for cost-effective live streaming speech synthesis evaluation. They identify verdict coupling, where multi-dimensional judges collapse dimension scores into a single preference bit, and propose Decoupled-Live-ProsodyJudge (D-LPJ) using masked SFT and a span-local GRPO strategy. Balanced-order LPJ beats a single Gemini call in point accuracy, and in a Best-of-8 TTS selection tournament the chosen utterance falls in the human top-3 for 85.29% of high-confidence cases.

arXiv cs.AI / cs.LG / cs.CL · 1d agoAI research

Closed-World Resolution Against Tool Hallucination in LLM Agents

Benchmark across ten LLMs documents 322 tool hallucinations and 154 more on MCP, showing model scale does not help and gates cannot reject fabricated calls.

The paper presents a five-class taxonomy (H1-H5) of tool hallucination in LLM agents, where models call nonexistent tools or pass arguments no schema declares, a blind spot no gating defense can reject since no gate made the decision. Across ten hosted models on two invocation surfaces, researchers measured 322 genuine hallucinations, concentrated on the unconstrained raw-JSON surface (34 versus 3), with model scale offering no benefit as a 675B model matched a 7-8B one. On the Model Context Protocol, merging servers into one namespace produced 154 hallucinations, including from frontier models that were clean on the single-registry surface. The versioned Hallucinated-Tools Benchmark (HTB) is released for comparable resolver evaluation.

arXiv cs.CR · 1d agoAI safety & security 2 sources

macOS 27 Golden Gate – Review

Ars Technica reviews macOS 27 Golden Gate, highlighting an unavoidable Apple Intelligence upgrade, new AFM 3 Core models, and dropped Intel Mac support.

macOS 27 Golden Gate delivers the first significant Apple Intelligence upgrade two years after launch, and the toggle to disable the AI features or delete downloaded models is gone. Apple Intelligence runs on a new AFM 3 Core model built in collaboration with Google, while the more capable AFM 3 Core Advanced requires an M3 chip and at least 12GB of RAM. The release drops all Intel Mac support, requiring Apple Silicon, with Sequoia security updates expected to end in fall 2027 and Tahoe's in 2028.

Prepared Or Unprepared? Evaluating Healthcare Workforce Readiness for Clinical Adoption of Artificial Intelligence in Nigeria

Survey of 761 Nigerian healthcare professionals finds high AI awareness (92.6%) but limited knowledge, preparedness, and major training and infrastructure barriers.

A cross-sectional study of 761 healthcare professionals across Nigeria, conducted from December 2025 to March 2026, found 92.6% awareness of AI in healthcare but 40.9% reporting low knowledge and only 63.0% feeling adequately prepared. Top barriers were lack of training (84.7%), poor infrastructure (71.1%), and high tool costs (61.0%). Willingness to adopt was strong, with 92.5% interested in training and 78.7% supporting AI in undergraduate curricula; preparedness differed significantly across geopolitical zones and professions.

arXiv cs.AI / cs.LG / cs.CL · 1d agoAI research

CaMeLoT: CaMeL orchestrated with Temporal logic for static verification and liveness

Researchers present CaMeLoT, extending CaMeL with CTL model checking that statically rejects unsafe LLM agent plans before any tool executes.

CaMeLoT adds a static verification layer to CaMeL, a runtime defense against prompt injection in tool-using LLM agents. It translates a generated plan into a finite-state transition system, labels it with tool calls, provenance, and taint information, and checks it against CTL temporal policies using the nuXmv model checker before any tool is invoked. Failed checks return counterexamples for plan repair, avoiding LLM calls, tool calls, and sandbox teardown. Evaluation covers policies derived from AgentDojo, SOC workflows, and prompt-extraction experiments.

arXiv cs.CR · 2d agoAI safety & security1

Amazon launches Alexa+ in India with Hindi support

Amazon launched its generative AI Alexa+ assistant in India with Hindi support in Early Access, free for Prime customers after testing.

Amazon announced that Alexa+, its generative AI-powered conversational assistant, is now available in India in Early Access with Hindi and English support, including mid-sentence language switching and long-form context retention. The assistant handles multi-step tasks such as ordering groceries via Amazon Now and controlling smart home devices, with integrations including Swiggy, District, MakeMyTrip, EazyDiner, Amazon Music, and JioSaavn. It will be free for Prime members after the testing period and cost about $20.85 per month for non-Prime customers. Amazon is targeting India's 600 million-plus Hindi speakers, and says smart device adoption grew 20% year over year.

TechCrunch · AI · 2d agoAI industry1

Inside NVIDIA’s cuDNN Graph API: Fusion, Autotuning, and Plan Reuse with cuDNN Frontend

MarkTechPost tutorial walks through NVIDIA's cuDNN Frontend graph API, covering kernel fusion, autotuning, plan reuse, and CUDA graph capture on Colab GPUs.

The tutorial explains how to express GPU computations as operation graphs via the cuDNN Frontend graph API, running the five-step build pipeline of validate, build operation graph, create execution plans, check support, and build plans. It progresses from a single fused convolution with bias and ReLU to autotuning across engine configs, FP8-style epilogues, attention, plan serialization, dynamic shapes, and CUDA graph capture. Each kernel is benchmarked against a PyTorch reference on a single Colab GPU to verify correctness and measure cost. The piece also covers practical setup issues like making libcudnn.so visible to the frontend's dynamic loader.

MarkTechPost · 2d agoAI tools & infra2

Building AI to accelerate science and improve lives

Google highlights AI-for-science advances: AlphaGenome Atlas mapping 9 billion genetic variants, WeatherNext 3 weather model, and global health AI tools.

Google detailed AI advances across science and health, including AlphaGenome Atlas, which mapped all 9 billion possible single-letter genetic changes in the human genome and was made openly available. WeatherNext 3 delivers 50% more accurate precipitation forecasts a day or more ahead and is already in products. AlphaFold is used by 4 million researchers in 190 countries, TB chest X-ray screening has processed 25,000+ scans across six nations, and the diabetic retinopathy model has supported 1.15 million screenings. Google also released its AI & Economy ATLAS global usage insights.

Google · AI · 2d agoAI industry

AI for everyone in every language

Google says its AI now spans 300+ languages reaching 7 billion people, unveiling Gemini 3.5 Transcribe, Live Translate, and TranslateGemma models.

Google announced its technologies now support more than 300 languages spoken by 7 billion people, 86% of the global population. Gemini 3.5 Live Translate powers real-time spoken translation across 70 languages and 2,000+ language pairs, while Gemini 3.5 Transcribe is its most precise speech-to-text model. Its Universal Speech Model was trained on 12 million hours of audio using cross-lingual transfer learning, and TranslateGemma is a family of lightweight open translation models covering 55 languages that run on-device. Open-data partnerships include WAXAL covering 27 Sub-Saharan African languages and Project Vaani with 30,000+ hours of speech across 109 languages.

Google · AI · 2d agoAI industry