Archive
207 posts
-
Open-weight AI is having its Kubernetes moment
-
As US weighs response to Chinese AI, industry urges against broad open-weight restrictions
-
Nvidia, Microsoft, Meta warn against overregulating open-weight models
-
Show HN: Echo – Fable-level results at 1/3 the cost using open-weight models
-
Runway launches AI model router as generative media gets crowded
-
AI chip startup Etched defies skeptics, hits $10.3B valuation from big-name investors
-
Show HN: I simulated closing the Strait of Hormuz on real oil trade data
-
Bringing Nunchaku 4-bit Diffusion Inference to Diffusers
-
Copilot vs. raw API access: What are you actually paying for?
-
Building AI infrastructure with the Effingham County community
-
Synthesia’s AI training platform is moving beyond videos into live coaching
-
Introducing OpenAI Presence
-
Introducing the ChatGPT for small business program
-
Kimi: Threat or menace?
-
Fine-tune video and image models at scale with NVIDIA NeMo Automodel and 🤗 Diffusers
-
Introducing Gemini 3.5 Flash Cyber
-
Why the first GPU financiers are turning to inference chips in a $400 million deal
-
A scorecard for the AI age
-
The AI compute gap: Enterprises are buying infrastructure faster than they can measure what it costs
-
The agent security gap: 54% of enterprises have already had an AI agent incident, and most still let agents share credentials
-
Roblox launches an AI-powered game-creation feature in its mobile app
-
The AI context gap: Enterprise AI organizations have a trust problem, not a retrieval problem — and most are still building the fix
-
The agent evaluation gap: Enterprise AI organizations have a reality-alignment problem, not a coverage problem — and most are shipping to production anyway
-
Agentic orchestration: Enterprise AI organizations have a deployment problem, not a platform problem — and most are calling chatbots agents
-
The US is advancing AI safety through state and federal action
-
Introducing Real World VoiceEQ: Measuring the human quality of voice AI
-
Empowering India’s next generation of innovators with ATL Saathi
-
Ollama: all aboard open models
-
Separating signal from noise in coding evaluations
-
Introducing GPT-Live
-
Expanding Managed Agents in Gemini API: background tasks, remote MCP and more
-
The latest AI news we announced in June 2026
-
ScarfBench: Benchmarking AI Agents for Enterprise Java Framework Migration
-
Introducing GeneBench-Pro
-
DiScoFormer: One transformer for density and score, across distributions
-
Faster Gemma 4 on MLX with multi-token prediction
-
HP Inc. launches Frontier strategic partnership with OpenAI
-
Accelerating Transformers Fine-Tuning with NVIDIA NeMo AutoModel
-
OpenAI and Broadcom unveil LLM-optimized inference chip
-
Introducing the FFASR Leaderboard: Benchmarking ASR in the Real World
-
Experimenting with the proposed Cross-Origin Storage API in Transformers.js
-
Daybreak: Tools for securing every organization in the world
-
Patch the Planet: a Daybreak initiative to support open source maintainers
-
New usage analytics and updated spend controls for enterprises
-
Beyond LoRA: Can you beat the most popular fine-tuning technique?
-
Introducing LifeSciBench
-
Predicting model behavior before release by simulating deployment
-
Introducing the OpenAI Partner Network
-
New OpenAI Academy courses for the next era of work
-
Introducing Gemma 4 12B: a unified, encoder-free multimodal model
-
Introducing the OpenAI Economic Research Exchange
-
The latest AI news we announced in May 2026
-
Improved performance and model support with GGUF
-
Nemotron 3.5 Content Safety: Customizable Multimodal Safety for Global Enterprise AI
-
Dreaming: Better memory for a more helpful ChatGPT
-
Introducing new capabilities to GPT-Rosalind
-
A blueprint for democratic governance of frontier AI
-
Introducing Mellum2: A 12B Mixture-of-Experts Model by JetBrains
-
OpenAI frontier models and Codex are now available on AWS
-
Strengthening societal resilience with Rosalind Biodefense
-
OpenAI’s Frontier Governance Framework
-
OpenJarvis: a local-first personal AI is now available to run with Ollama
-
How Virgin Atlantic ships faster with Codex
-
Introducing OpenAI for Singapore
-
Google just redesigned the search box for the first time in 25 years — here’s why it matters more than you think.
-
Introducing the Ettin Reranker Family
-
Simulate real-world places with Project Genie and Street View
-
Databricks brings GPT-5.5 to enterprise agent workflows
-
Co-Scientist: A multi-agent AI partner to accelerate research
-
What Parameter Golf taught us about AI-assisted research
-
Building Blocks for Foundation Model Training and Inference on AWS
-
OpenAI launches DeployCo to help businesses build around intelligence
-
Advancing voice intelligence with new models in the API
-
Introducing Trusted Contact in ChatGPT
-
Introducing ChatGPT Futures: Class of 2026
-
Unlocking large scale AI training networks with MRC (Multipath Reliable Connection)
-
Enabling a new model for healthcare with AI co-clinician
-
Introducing Advanced Account Security
-
DeepInfra on Hugging Face Inference Providers 🔥
-
Introducing NVIDIA Nemotron 3 Nano Omni: Long-Context Multimodal Intelligence for Documents, Audio and Video Agents
-
OpenAI models, Codex, and Managed Agents come to AWS
-
OpenAI available at FedRAMP Moderate
-
Modeling an AI jobs transition
-
Introducing GPT-5.5
-
Speeding up agentic workflows with WebSockets in the Responses API
-
Introducing workspace agents in ChatGPT
-
Introducing OpenAI Privacy Filter
-
Introducing ChatGPT Images 2.0
-
QIMMA قِمّة ⛰: A Quality-First Arabic LLM Leaderboard
-
Scaling Codex to enterprises worldwide
-
Introducing GPT-Rosalind for life sciences research
-
Accelerating the cyber defense ecosystem that protects us all
-
Training and Finetuning Multimodal Embedding & Reranker Models with Sentence Transformers
-
Gemini 3.1 Flash TTS: the next generation of expressive AI speech
-
The next evolution of the Agents SDK
-
Trusted access for the next era of cyber defense
-
Our response to the Axios developer tool compromise
-
Multimodal Embedding & Reranker Models with Sentence Transformers
-
Introducing the Child Safety Blueprint
-
Welcome Gemma 4: Frontier multimodal intelligence on device
-
Granite 4.0 3B Vision: Compact Multimodal Intelligence for Enterprise Documents
-
Ollama is now powered by MLX on Apple Silicon in preview
-
Inside our approach to the Model Spec
-
Introducing the OpenAI Safety Bug Bounty program
-
Helping developers build safer AI experiences for teens
-
Update on the OpenAI Foundation
-
Powering product discovery in ChatGPT
-
A New Framework for Evaluating Voice Agents (EVA)
-
Measuring progress toward AGI: A cognitive framework
-
Introducing GPT-5.4 mini and nano
-
OpenAI Japan announces Japan Teen Safety Blueprint to put teen safety first
-
From model to agent: Equipping the Responses API with a computer environment
-
New ways to learn math and science in ChatGPT
-
Introducing Storage Buckets on the Hugging Face Hub
-
Introducing GPT-5.4
-
Introducing ChatGPT for Excel and new financial data integrations
-
Introducing the Adoption news channel
-
Introducing Modular Diffusers - Composable Building Blocks for Diffusion Pipelines
-
Understanding AI and learning outcomes
-
Introducing the Stateful Runtime Environment for Agents in Amazon Bedrock
-
Pacific Northwest National Laboratory and OpenAI partner to accelerate federal permitting
-
OpenAI announces Frontier Alliance Partners
-
Introducing OpenAI for India
-
Introducing EVMbench
-
Introducing Lockdown Mode and Elevated Risk labels in ChatGPT
-
Introducing GPT-5.3-Codex-Spark
-
Bringing ChatGPT to GenAI.mil
-
Transformers.js v4: Now Available on NPM!
-
Introducing SyGra Studio
-
Introducing Trusted Access for Cyber
-
Introducing OpenAI Frontier
-
Introducing GPT-5.3-Codex
-
Unlocking the Codex harness: how we built the App Server
-
Introducing the Codex app
-
Retiring GPT-4o, GPT-4.1, GPT-4.1 mini, and OpenAI o4-mini in ChatGPT
-
Introducing Daggr: Chain apps programmatically, inspect visually
-
The next chapter for AI in the EU
-
Introducing Prism
-
Unrolling the Codex agent loop
-
Railway secures $100 million to challenge AWS with AI-native cloud infrastructure
-
Introducing Edu for Countries
-
Differential Transformer V2
-
Our approach to age prediction
-
Introducing Waypoint-1: Real-time interactive video diffusion from Overworld
-
Claude Code costs up to $200 a month. Goose does the same thing for free.
-
A business that scales with the value of intelligence
-
Introducing ChatGPT Go, now available worldwide
-
Claude Code with Anthropic API compatibility
-
Strengthening the U.S. AI supply chain through domestic manufacturing
-
OpenAI Codex with Ollama
-
OpenAI partners with Cerebras
-
Zenken boosts a lean sales team with ChatGPT Enterprise
-
Salesforce rolls out new Slackbot AI agent as it battles Microsoft and Google in workplace AI
-
Anthropic launches Cowork, a Claude Desktop agent that works in your files — no coding required
-
Nous Research's NousCoder-14B is an open-source coding model landing right in the Claude Code moment
-
Introducing ChatGPT Health
-
Introducing Falcon-H1-Arabic: Pushing the Boundaries of Arabic Language AI with Hybrid Architecture
-
Announcing OpenAI Grove Cohort 2
-
AprielGuard: A Guardrail for Safety and Adversarial Robustness in Modern LLM Systems
-
Evaluating chain-of-thought monitorability
-
Deepening our collaboration with the U.S. Department of Energy
-
Introducing GPT-5.2-Codex
-
Introducing OpenAI Academy for News Organizations
-
Developers can now submit apps to ChatGPT
-
Gemma Scope 2: helping the AI safety community deepen understanding of complex language model behavior
-
Evaluating AI’s ability to perform scientific research tasks
-
Measuring AI’s capability to accelerate biological research
-
The new ChatGPT Images is here
-
BBVA and OpenAI collaborate to transform global banking
-
Introducing GPT-5.2
-
The Walt Disney Company and OpenAI reach landmark agreement to bring beloved characters to Sora
-
FACTS Benchmark Suite: Systematically evaluating the factuality of large language models
-
Introducing swift-huggingface: The Complete Swift Client for Hugging Face
-
Introducing OpenAI for Australia
-
We Got Claude to Fine-Tune an Open Source LLM
-
Announcing the initial People-First AI Fund grantees
-
Mixpanel security incident: what OpenAI users need to know
-
Expanding data residency access to business customers worldwide
-
OVHcloud on Hugging Face Inference Providers 🔥
-
Introducing shopping research in ChatGPT
-
20x Faster TRL Fine-tuning with RapidFire AI
-
Early experiments in accelerating science with GPT-5
-
Introducing AnyLanguageModel: One API for Local and Remote LLMs on Apple Platforms
-
Building more with GPT-5.1-Codex-Max
-
Introducing OpenAI for Ireland
-
SIMA 2: An Agent that Plays, Reasons, and Learns With You in Virtual 3D Worlds
-
Introducing group chats in ChatGPT
-
Introducing GPT-5.1 for developers
-
GPT-5.1: A smarter, more conversational ChatGPT
-
Introducing the Teen Safety Blueprint
-
Introducing IndQA
-
Introducing Aardvark: OpenAI’s agentic security researcher
-
Introducing gpt-oss-safeguard
-
gpt-oss-safeguard technical report
-
Doppel’s AI defense system stops attacks before they spread
-
MiniMax M2
-
MedGemma: Our most capable open models for health AI development
-
Gemini 2.5 Flash-Lite is now ready for scaled production use
-
Aeneas transforms how historians connect the past
-
Strengthening our Frontier Safety Framework
-
Introducing CodeMender: an AI agent for code security
-
Try Deep Think in the Gemini app
-
Introducing Gemma 3 270M: The compact model for hyper-efficient AI
-
VaultGemma: The world's most capable differentially private LLM
-
Consensus accelerates research with GPT-5 and Responses API
-
Building the Open Agent Ecosystem Together: Introducing OpenEnv
-
The next chapter for UK sovereign AI