LIVE PULSE
4.9 Anthropic CEO Amodei calls for slower AI development and shared safety rules11 src2.6 Agility Robotics unveils Digit 5 humanoid for warehouses and factories2 src2.4 Apple ships rebuilt Siri with Google Gemini, but not in the EU2 src2.2 Siri AI in macOS 27 Golden Gate: FAQ, Germany availability, privacy questions2 src1.7 Sam Altman says OpenAI will not go public in 2026, citing AI safety concerns5 src1.4 OpenAI contractors review real ChatGPT conversations to rate responses, report says2 src1.3 Anthropic data retention policy prompts firms to limit Claude use for sensitive work1 src1.3 CoMem Paper Proposes Shared and Individual Memory Design for LLM Multi-Agent Systems1 src1.3 Paper proposes evolving context parameterization for large language models1 src1.3 Fine-Tuning Vision-Language Models with Listener Gaze for Referring Expressions1 src4.9 Anthropic CEO Amodei calls for slower AI development and shared safety rules11 src2.6 Agility Robotics unveils Digit 5 humanoid for warehouses and factories2 src2.4 Apple ships rebuilt Siri with Google Gemini, but not in the EU2 src2.2 Siri AI in macOS 27 Golden Gate: FAQ, Germany availability, privacy questions2 src1.7 Sam Altman says OpenAI will not go public in 2026, citing AI safety concerns5 src1.4 OpenAI contractors review real ChatGPT conversations to rate responses, report says2 src1.3 Anthropic data retention policy prompts firms to limit Claude use for sensitive work1 src1.3 CoMem Paper Proposes Shared and Individual Memory Design for LLM Multi-Agent Systems1 src1.3 Paper proposes evolving context parameterization for large language models1 src1.3 Fine-Tuning Vision-Language Models with Listener Gaze for Referring Expressions1 src
HEATPULSEAI MAGAZINES
FLIP · FOLLOW · SAVE

Speech LLMs

topic3 events
papersTODAY 04:00 UTC

SURE-Voice: Training-Free Front End Filters Speech Evidence for Speech LLMs

A new arXiv paper introduces SURE-Voice, a training-free front-end component that estimates whether audio actually contains usable speech evidence before a speech language model generates output. The work frames the problem as pre-generation support estimation, targeting cases where speech LLMs produce plausible but unsupported responses from audio lacking meaningful speech.

papersSEP 12 04:00 UTC

RetroThinker adds retrospective thinking to speech LLMs

A new arXiv paper introduces RetroThinker, a method that lets speech large language models revisit and revise earlier reasoning steps. Speech LLMs cut latency and preserve vocal cues lost in pipelines that combine speech recognition with text models, but they still trail text-only systems in capability. The work aims to narrow that gap by adapting retrospective reasoning to the speech setting.

papersSEP 12 04:00 UTC

Logit-Space Integration Improves Contextual Biasing in Speech LLMs

A new arXiv paper proposes LOGIC, a method that injects contextual biasing into Speech LLMs within logit space rather than through prompting. The approach targets the difficulty these models have in recognizing newly emerging entities such as names, trends, and personalized terms. The authors report gains in both efficiency and robustness compared with prompt-based baselines.