GPT-5.1 Instant and GPT-5.1 Thinking System Card Addendum
OpenAI published a system card addendum for GPT-5.1 Instant and Thinking, covering updated safety evals including mental health and emotional reliance.
Use this view to inspect the underlying evidence corpus. For ranked developments, decision posture and interpretation, use Signals.
Raw feed or Signals?
Raw feed is chronological evidence. Signals ranks and interprets material change.
OpenAI published a system card addendum for GPT-5.1 Instant and Thinking, covering updated safety evals including mental health and emotional reliance.
OpenAI releases GPT-5.1, a GPT-5 series upgrade with improved conversational tone and user-facing customization options.
Google DeepMind research details how AI visual perception differs from human perception, impacting object recognition and scene understanding.
Anthropic secured $13 billion in funding, with commentary suggesting the investment emphasizes AI safety and potential future regulation.
Google DeepMind pilot in Northern Ireland schools with Gemini and other generative AI tools saved teachers 10 hours weekly.
OpenAI claims new guardrails for ChatGPT reduce misinformation and harmful content risks, improving trust in the platform.
Research on deformable object dynamics, efficient inference, and reliable simulation indicates advances in modeling complex physical interactions for robotics and AI.
Anthropic's reported payout in legal dispute highlights growing pressure on AI developers regarding creator rights and copyright. Broader implications for model training data use.
Report claims OpenAI quietly acquired Statsig, a platform for product experimentation and feature flagging, potentially integrating A/B testing into model development.
OpenAI publishes explainer on prompt injection attacks, covering attack mechanics and its mitigation research and safeguards.
Notion rebuilt its AI layer on GPT-5 to enable autonomous, multi-step agents in Notion 3.0 productivity workflows.
BBVA reports 20,000+ custom GPTs built and claimed efficiency gains up to 80% after deploying ChatGPT Enterprise org-wide.
Chime CMO describes shift to AI-driven, agent-based marketing model and advocates for AI literacy among marketing leaders.
OpenAI reports crossing 1 million business customers globally across healthcare, financial services, and other sectors.
Expert commentary on trust implications of Grok leaks, discussing psychological, technological, and policy angles around AI adoption.
Nvidia's market valuation reaches $5 trillion, claiming a dominant position in the AI ecosystem beyond just hardware.
OpenAI launches IndQA, a benchmark for AI evaluation across 12 Indian languages and 10 knowledge domains, built with domain experts.
ChatGPT can now connect to internal company data systems, allowing it to read reports and generate insights from proprietary files.
OpenAI and AWS sign multi-year, $38B partnership for AWS to provide compute infrastructure for OpenAI model training and deployment.
OpenAI announced the acquisition of Sky, a move to enable AI to autonomously handle digital tasks, focusing on human-AI collaboration.
Report claims OpenAI is developing 'Atlas', an AI-powered browser, marketed as an 'AI companion' for internet exploration.
General Intuition secured $134M to develop AI with human-like spatial reasoning, aiming for next-generation AI cognition.
Google's Android details how it uses AI to protect users from mobile scams, noting a $400 billion global loss from AI-driven fraud.
OpenAI launches Aardvark, an autonomous AI security researcher in private beta that finds, validates, and helps remediate software vulnerabilities.
The podcast 'No Priors' discusses how the ElevenLabs API is being used to innovate in music composition and sound design.
OpenAI published technical details on OWL, the browser architecture powering ChatGPT Atlas, decoupling Chromium for agentic web browsing.
Google DeepMind launches 'AI for Math Initiative' with research institutions to apply AI in advanced mathematical discovery.
NVIDIA Isaac Sim framework used to develop and deploy healthcare robots, highlighting simulation-to-real-world transfer for complex automation.
OpenAI releases gpt-oss-safeguard 120B and 20B: open-weight models trained to classify content against a provided policy.
OpenAI releases gpt-oss-safeguard, open-weight reasoning models for safety classification with customisable policy enforcement.
Doppel deploys GPT-5 with reinforcement fine-tuning to detect deepfake/impersonation attacks, claiming 80% analyst workload reduction.
METR Research reviewed Anthropic's Summer 2025 Pilot Sabotage Risk Report, which assesses "sabotage risk" from Claude Opus 4 and 4.1 as low but non-negligible.
Microsoft and OpenAI sign updated partnership agreement expanding long-term collaboration and responsible AI commitments.
OpenAI announces recapitalization and governance restructuring, framing it as mission-aligned expansion of resources for responsible AI.
Jack Clark's Import AI #433 covers AI auditors, robotics research, and AI-managed laboratory automation.
OpenAI worked with 170+ mental health experts to improve ChatGPT's distress recognition, claiming 80% reduction in unsafe responses.
OpenAI published a GPT-5 system card addendum detailing safety benchmarks for emotional reliance, mental health, and jailbreak resistance.
Hugging Face released huggingface_hub v1.0, marking five years of their open-source machine learning platform and tools.
Datumo is presented as a new competitor to Scale AI, potentially offering faster, more cost-effective AI training data services.
Google DeepMind's Gemini 2.5 Flash-Lite, a cost-efficient model with a 1 million-token context window and multimodality, is now generally available.
Discussion on OpenArt and 'brainrot content' in AI-generated media, covering the promises and perils of algorithm-driven art platforms.
Google DeepMind's AlphaEarth Foundations integrates petabytes of Earth observation data into a unified representation for global mapping.
Mistral AI launched Mistral AI Studio, a platform offering direct API access to their models, including new enterprise-focused features.
Google DeepMind's experimental Backstory AI tool aids users in tracing the context and origin of online images.
Google DeepMind's advanced Gemini model, Deep Think, achieved a gold-medal standard on the International Mathematical Olympiad (IMO) problems.
Google DeepMind's new Perch model analyzes bioacoustic data to monitor endangered species, accelerating conservation efforts.
Google DeepMind's Gemini 2.5 Deep Think achieved gold-medal level in the International Collegiate Programming Contest, demonstrating advanced abstract problem-solving.
Nvidia is reportedly developing 'Cosmos,' a world model aiming to create smarter AI. Details are scarce, suggesting a research-stage initiative.
Google DeepMind announced an update to its Frontier Safety Framework to identify and mitigate severe risks from advanced AI models.
Google DeepMind's Gemini Robotics 1.5 enables robots with advanced perception, planning, tool use, and multi-step task execution.
© 2026 OneBench: AI Insights. All rights reserved.
Evidence before opinion