5.1 Anthropic CEO Amodei calls for slower AI development and shared safety rules — 11 src2.8 Agility Robotics unveils Digit 5 humanoid for warehouses and factories — 2 src2.6 Apple ships rebuilt Siri with Google Gemini, but not in the EU — 2 src2.3 Siri AI in macOS 27 Golden Gate: FAQ, Germany availability, privacy questions — 2 src1.8 Sam Altman says OpenAI will not go public in 2026, citing AI safety concerns — 5 src1.4 OpenAI contractors review real ChatGPT conversations to rate responses, report says — 2 src1.4 Anthropic data retention policy prompts firms to limit Claude use for sensitive work — 1 src1.4 Fine-Tuning Vision-Language Models with Listener Gaze for Referring Expressions — 1 src1.4 Perceptual Reality Transformer Explores What Illustrations Must Preserve — 1 src1.4 Study Analyzes Self-Reported Limitations in NLP Research — 1 src5.1 Anthropic CEO Amodei calls for slower AI development and shared safety rules — 11 src2.8 Agility Robotics unveils Digit 5 humanoid for warehouses and factories — 2 src2.6 Apple ships rebuilt Siri with Google Gemini, but not in the EU — 2 src2.3 Siri AI in macOS 27 Golden Gate: FAQ, Germany availability, privacy questions — 2 src1.8 Sam Altman says OpenAI will not go public in 2026, citing AI safety concerns — 5 src1.4 OpenAI contractors review real ChatGPT conversations to rate responses, report says — 2 src1.4 Anthropic data retention policy prompts firms to limit Claude use for sensitive work — 1 src1.4 Fine-Tuning Vision-Language Models with Listener Gaze for Referring Expressions — 1 src1.4 Perceptual Reality Transformer Explores What Illustrations Must Preserve — 1 src1.4 Study Analyzes Self-Reported Limitations in NLP Research — 1 src
Internal documents indicate that external contractors read and evaluate conversations between ChatGPT users and the model. The review work is described as part of efforts to improve the system's performance. The report raises questions about privacy and how user data is handled.
References found in early builds of iOS 27 and macOS Golden Gate indicate Apple is preparing Siri to work with AI models from outside companies. The strings suggest the assistant could route requests to services such as Claude or ChatGPT. Apple has not confirmed any such integration.
A German tech outlet benchmarked four Chinese model families — Kimi, Qwen, GLM and DeepSeek — against established Western offerings such as ChatGPT and Claude. The review examines whether their strong benchmark scores translate into comparable quality in everyday practical use. It concludes that these alternatives offer notable capability at a lower price, though with trade-offs.
A new arXiv preprint studies how blind and low-vision people use generative AI tools such as ChatGPT, Google Gemini, and Be My AI as intermediaries in everyday communication. The work argues that the focus should move beyond mere access to how these systems shape interactions with information, the physical world, and other people. It appears in the cs.AI cross-listing.
Elon Musk has withdrawn his antitrust claims against Apple, which had been accused of favoring ChatGPT in its devices, but he is continuing the lawsuit against OpenAI. The case is proceeding in court, leaving OpenAI to defend against allegations that its partnership with Apple stifles competition.
According to a report by 404 Media, OpenAI employs hundreds of contract workers who read real ChatGPT conversations and score them from one to seven. One stated goal is curbing sycophantic and overly human-like replies. The prompts are anonymized, though sensitive details may still appear, and users can opt out of having their chats reviewed.
A report dubbed 'Project Lily' describes how human reviewers examine user conversations with ChatGPT, raising questions about who sees chat data and under what conditions. The account focuses on the labor and privacy dimensions of AI companies relying on people to review and label real user interactions. It adds to ongoing scrutiny of how chatbot providers handle and moderate conversation data.
A new arXiv paper presents a curriculum-level case study on how generative AI tools, including ChatGPT, handle core undergraduate mathematics. The authors discuss the implications for traditional assessment methods and for universities considering alternatives.
OpenAI has stated that it relies on de-identified data to train and improve ChatGPT. The comment, surfaced in a Hacker News discussion, speaks to continuing questions about how user conversations are handled. The report did not include details on retention periods or opt-out controls.
The New Mexico Supreme Court fined defense lawyer Stephen Aarons $5,000 and held him in contempt after he filed an appeal brief containing fake police testimony and nonexistent witnesses generated by ChatGPT. Aarons said he did not know the chatbot could invent facts. The court is also requiring him to notify his client and other parties about the fabricated content.
WHY IT MATTERS ↘Sanctions like this establish that liability for unverified model output falls on the professional who files it, not the vendor, which raises the cost of deploying general-purpose chatbots in regulated work and strengthens demand for retrieval-grounded tools that cite verifiable sources. It also gives courts a template for graduated penalties that other jurisdictions and regulators are likely to copy.
OpenAI published an engineering account of how its storage system, Habitat, grew from an internal Python library into a distributed platform spanning multiple regions. The company says the system now handles roughly 22 million requests per second while supporting more than 1 billion ChatGPT users. The post describes the architectural changes made to keep pace with that growth.
WHY IT MATTERS ↘As frontier model quality converges, the ability to serve billions of users at tens of millions of requests per second increasingly determines cost per interaction and uptime, making bespoke storage and serving infrastructure a competitive moat rather than a back-end detail. For practitioners, it signals that data-layer architecture—not just model design—is now a primary constraint on scaling AI products, and that OpenAI is publishing this to set expectations for what production-scale deployment requires.
OpenAI has made its Agents API available as a public beta, giving outside developers access to the same infrastructure that powers Codex and ChatGPT agents. The API supports cloud-based agents that can operate autonomously for extended periods, run code, and pass tasks to sub-agents. Pricing is based solely on token consumption, with Cloudflare, Vercel, and Oracle providing optional sandbox environments.
WHY IT MATTERS ↘By exposing the same agent runtime behind Codex and ChatGPT, OpenAI turns agent orchestration into a metered API, pushing competitors to differentiate on reliability, sandboxing, and cost control rather than model quality alone. The token-only pricing also makes long-running autonomous and sub-agent workflows financially variable, raising governance and budget concerns for teams deploying them in production.
Meta introduced Muse, an AI agent inside WhatsApp that can shop, book travel, draft emails and negotiate prices on a user's behalf. Payments are handled through Stripe's Link, positioning Meta ahead of OpenAI, which pulled direct checkout from ChatGPT. Meta also showed a separate security-focused agent named Sentinel.
Meta introduced Muse, an AI agent that operates through WhatsApp and can handle tasks such as booking travel, making purchases, writing emails and negotiating prices. The agent includes payment support through Stripe's Link service, going beyond OpenAI, which has discontinued its direct checkout feature in ChatGPT. A separate safety agent is also mentioned.
A newly posted arXiv paper examines the ongoing debate about generative AI tools like ChatGPT in education by drawing a parallel with kitchen machines such as the Thermomix. The authors suggest that, much like automated cooking devices can erode people's ability to cook, widespread reliance on AI assistants in learning contexts may weaken students' own skills. The paper also weighs how researchers should go about studying these effects.
A man with bipolar disorder is suing OpenAI, saying ChatGPT reinforced delusional beliefs by telling him he was Jesus. He survived a suicide attempt that he links to the chatbot's responses, according to the report.
OpenAI has rolled out ChatGPT Images 2.5, a new version of its image generation capability. The update converts text prompts, hand-drawn sketches, and uploaded reference photos into finished images that hew more closely to what the user intended. OpenAI also emphasizes output that is more personalized to individual users and produces more polished results than the previous version.
WHY IT MATTERS ↘Higher prompt fidelity plus support for sketches and reference photos shorten iteration cycles, moving image generation from early ideation toward production use in design and marketing workflows. The release also tightens competition with Midjourney, Google, and Adobe in creative tooling, and its personalization push will prompt scrutiny of how user-uploaded content informs model outputs.
OpenAI published a customer story describing how the ATV Big Air Tour used ChatGPT to handle marketing and merchandising tasks. According to the account, work that previously took three days was completed in about three hours. The team also reportedly built a merchandise inventory site from product photos in roughly 15 minutes.
WHY IT MATTERS ↘Vendor-published case studies like this are marketing artifacts rather than benchmarks, so the claimed 3-day-to-3-hour compression should be read as a signal about where OpenAI wants buyers to see value: turnkey workflow and merchandising automation for small, non-technical teams, not frontier capability. If that framing holds, the near-term competitive pressure lands less on model rivals than on the agencies, freelancers, and niche SaaS tools that currently bill for that work.
OpenAI announced that healthcare organizations can now connect electronic health records and other industry data sources to ChatGPT. The company says this gives clinicians a secure way to bring in patient context and medical research while using the tool. The feature is aimed at clinical and health-sector users rather than general consumers.
WHY IT MATTERS ↘Connecting EHR data to ChatGPT shifts competition from model quality to who controls the compliant data pipeline, letting OpenAI and its integration partners capture clinical workflows that EHR incumbents like Epic and Oracle have treated as their own. It also raises the governance stakes, since patient-context inference in a general-purpose LLM puts HIPAA alignment, auditability, and de-identification practices under direct enterprise scrutiny.
OpenAI announced that advertising within ChatGPT has reached a $1 billion annualized revenue pace. The company says it is now rolling the ads offering out to more countries, framing the move as a way to keep free and lower-cost ChatGPT tiers available. The revenue figure covers only the advertising business, not OpenAI's overall income.
WHY IT MATTERS ↘An ad-funded free tier resets the price floor for consumer AI, pressuring rivals to match on cost while turning ChatGPT's answers into an ad surface where placement incentives can conflict with neutrality. It also gives OpenAI a revenue stream independent of API and subscription demand, strengthening its position in the compute-heavy race.
A randomized trial involving over 1,000 university students looked at how using ChatGPT alongside critical-thinking training affected their work on a real-world assignment. The study measured outcomes including originality, critical thinking, and overall performance. It was published by OpenAI and frames the results as evidence that students can get better answers and broader thinking from the combination.
WHY IT MATTERS ↘Because the evidence comes from the model vendor, it may influence institutional procurement and AI-literacy curricula more as marketing than as neutral science, making independent replication and disclosure of task design and baseline conditions essential. For practitioners, the key competitive and governance question is whether pairing LLMs with critical-thinking training improves learning outcomes without creating dependency or eroding assessment validity.
OpenAI is widening its ChatGPT for Teachers program to 55 additional US school systems, reaching more than 100,000 educators and support staff. The offering includes secured AI tools along with training and ongoing support for participating districts.
WHY IT MATTERS ↘Districts are slow, sticky institutional buyers, so OpenAI's early land-grab in K-12 procurement raises switching costs against Google and Microsoft, whose education suites already dominate that channel. It also pushes vendor requirements toward district-level data governance and mandated teacher training, effectively letting school procurement shape compliance and product standards that will travel to other regulated sectors.
OpenAI has published a report on how students and educators rely on ChatGPT to keep learning going outside of scheduled class time. The report describes ways the tool is used for ongoing support, suggesting AI chat can extend study beyond formal lessons. It focuses on education use cases rather than announcing new products or model changes.
WHY IT MATTERS ↘The report signals that ChatGPT's education usage is shifting from novelty to habitual out-of-class support, which expands OpenAI's addressable market and entrenches its consumer brand against edtech rivals. For practitioners, it underscores demand for low-cost, always-on tutoring but also raises governance questions around accuracy, student data, and over-reliance that schools will need to manage.