GPT-5.6 把复杂生产工作流推向更低 token 成本
GPT-5.6 的重点不是单一跑分,而是用更少输出 token 完成复杂生产任务,并以 sol / terra / luna 覆盖质量、成本与吞吐量梯度。
跨来源聚合,保留权重判断与完整时间线。
原文可溯源GPT-5.6 的重点不是单一跑分,而是用更少输出 token 完成复杂生产任务,并以 sol / terra / luna 覆盖质量、成本与吞吐量梯度。
可靠 Agent 的共同点是可组合、可观察的小步骤;只有当评测证明收益时,才增加规划、循环和多 Agent 协作。
Jalapeño’s first results show industry-leading speed and efficiency in AI inference。Jalapeño is a custom inference chip from OpenAI that delivers faster, more power-efficient AI inference, with higher throughput and lower latency for modern models.
“泡面盈利”不是成熟公司利润最大化,而是尽快覆盖创始人基本生活费,从而把融资从生死期限变成可选择的增长工具。
一人公司模型的核心是限制固定组织成本:产品、支持、支付和分发尽可能软件化,把专业劳动留给最有杠杆的环节。
EvolveTrade: Experience-Driven Policy Refinement for Self-Evolving LLM Trading Agents。arXiv:2609.17632v1 Announce Type: new Abstract: Large language model (LLM) trading agents can combine market data, news, and executable analysis, but their behavior is often controlled by static hand-written tool-use policies that are fixed before deployment.
ZGCM-1: A Fully Open and Extremely Efficient Foundation Model for Math and Agentic Search。arXiv:2609.13356v1 Announce Type: new Abstract: In this work, we present ZGCM-1, a fully open 7B dense foundation model trained from scratch with extreme data, system, and algorithmic efficiency.
Adaptive Entangled Game Modules in Artificial General Intelligence。arXiv:2609.09226v1 Announce Type: new Abstract: We introduce a probability-wave framework for modeling the collective behavior of interacting adaptive agents, deriving testable eigenmodes through a generalized behavioral intelligence (GBI) nonlocal probability-wave equation.
OpenDiscoveryTrace: Process Traces for Evaluating AI Scientist Workflows。arXiv:2609.09203v1 Announce Type: new Abstract: Existing benchmarks for autonomous AI scientists evaluate only final outputs---generated code, hypotheses, or papers---yet discard the reasoning process by which those outputs were obtained.
When Does Memory Help? A Cost-Aware Evaluation of Long-Term Memory in Tool-Using LLM Agents。arXiv:2609.05441v1 Announce Type: new Abstract: Long-term memory for LLM agents is evaluated today by conversational recall benchmarks (LoCoMo, LongMemEval), which measure question answering over dialogue history, not whether remembered facts change what a t
Harbor Adapters and Harbor-Index: Infrastructure and a Curated Meta-Dataset for Large-Scale Agentic Evaluation。arXiv:2609.04298v1 Announce Type: new Abstract: Evaluating agents on the growing number of agentic benchmarks is challenging because they often require complex environments and agent integrations.
From Matching Models to Recruiting Agents: A Systematized Narrative Review of AI Recruitment Systems, Evaluation, and Governance。arXiv:2609.04286v1 Announce Type: new Abstract: Artificial intelligence in recruitment has shifted the object being automated from profile pairs and ranked lists to multi-stage workflows that retrieve evidence, compare ca
EXAONE Forecast for Finance。arXiv:2609.04239v1 Announce Type: new Abstract: This technical report presents EXAONE Forecast for Finance (EXAONE Finance), a financial time series (TS) foundation model (TSFM) tailored to financial forecasting.
Speculative Macro Commit for Faster Tool-Using Agents。arXiv:2609.03236v1 Announce Type: new Abstract: Tool-using LLM agents spend wall-clock time not only on model inference but also in serial action--observation turns, where each tool call, environment transition, and observation can delay subsequent decisions.
EvalDetectBench: A Benchmark for Measuring Evaluation Awareness in Frontier Language Models。arXiv:2609.01611v1 Announce Type: new Abstract: Frontier large language models can often recognize when they are being evaluated, a capability known as evaluation awareness.
A collective capability boundary in frontier large language models on guideline-conformant and case-specific oncology decision-making。arXiv:2608.28592v1 Announce Type: new Abstract: Large language models (LLMs) achieve high scores on medical knowledge examinations, yet real-world oncology is not a knowledge test--it is a sequence of guideline-pathw
DS-Lighting: Making Agent Harnesses Explicit for Data-Science Automation。arXiv:2608.28590v1 Announce Type: new Abstract: Large Language Model (LLM) agents have shown promise for automating data-science workflows, yet their end-to-end performance depends critically on the agent harness that represents tasks, manages execution state, constrains outpu
Standalone LLM and a Pre-specified Agentic Pipeline for Explaining ICU Mortality Predictions: a Feasibility Study on the eICU Demo Dataset。arXiv:2608.26109v1 Announce Type: new Abstract: Machine-learning models can predict ICU mortality accurately, but feature-attribution methods alone rarely provide the clinical narrative needed for bedside use.
Nvidia revenue doubles on continued AI demand。Chipmaker's $96bn quarter beat forecasts as demand for AI hardware continues to accelerate.
LLM Agents Perform Controlled Experiments Using Simulation Models。arXiv:2608.23622v1 Announce Type: new Abstract: Large language models (LLMs) have shown strong capabilities in reasoning, planning, and tool use, but many scientific and engineering tasks require more than plausible text and code generation.
KVBoost: Chunk-Level Key-Value Cache Reuse with Deviation-Guided Recomputation for Efficient Large Language Model Inference。arXiv:2608.21362v1 Announce Type: new Abstract: Transformer-based large language models (LLMs) incur high prefill latency because key-value (KV) tensors must be recomputed for each request.
PrimeAgentOrchestrator: Memory-Primed Agent Spawning for Personal AI Infrastructure。arXiv:2608.20342v1 Announce Type: new Abstract: Large language model (LLM) coding agents start each session with an empty context window, discarding accumulated knowledge from prior work.
Active Inference as Context Acquisition for AI Agents。arXiv:2608.19202v1 Announce Type: new Abstract: Interactive AI agents must acquire the right context as efficiently as possible.
Large Language Models in Mental Health: A Systematic Review of Applications, Innovations, and Ethical Challenges。arXiv:2608.18080v1 Announce Type: new Abstract: We present a review on the applications of large language models (LLMs) in health, e.g., social media analysis, clinical conversational agents, therapy support tools, prompt engineering, mu
Position: Collusion Risks Among AI Reasoning Agents Justify Certification Requirements for Making Market Decisions。arXiv:2608.18078v1 Announce Type: new Abstract: This position paper argues that AI agents with chain-of-thought reasoning capabilities are predisposed to exhibit collusive behavior and should be required to obtain behavioral certificat
GxP-Agent: Process-DAG Topology for Reliable Clinical Trial Programming with LLM Agents。arXiv:2608.16890v1 Announce Type: new Abstract: Clinical trial programming -- transforming study protocols into analysis-ready datasets under CDISC standards -- is a bottleneck in regulatory submissions, yet LLM-based code generation fails catastrophically on th
The Unwritten Benchmark: A New Challenge for Multimodal Machine Learning in Abstract Perceptual Reasoning。arXiv:2608.14558v1 Announce Type: new Abstract: Current multimodal models have demonstrated remarkable proficiency in recognizing static visual and auditory content.
Large Language Models Show Metacognitive Sensitivity in Medical Reasoning。arXiv:2608.14552v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly evaluated and used in medicine, but clinical usefulness depends on answer accuracy and whether confidence tracks evidence quality and uncertainty.
Depth-Aware Sensitivity Analysis of Mixture-of-Experts Models via Magnitude-Based Expert Masking。arXiv:2608.13565v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) architectures scale large language models (LLMs) while preserving computational efficiency through sparse activation.
Diagnostic Foundation for Evaluating LLMs' Research Integrity as Co-Scientists。arXiv:2608.12345v1 Announce Type: new Abstract: Language models are increasingly deployed as co-scientists, yet their ability to uphold research integrity under institutional pressure remains unmeasured.
Distribird: Literature-Informed Prior Distribution Design for Bayesian Model Calibration。arXiv:2608.11210v1 Announce Type: new Abstract: Bayesian calibration of process-based models requires a prior distribution for each model parameter.
Dynamic Governance of Multi-LLM Agent Systems for Collaborative Conversational Outcomes。arXiv:2608.11207v1 Announce Type: new Abstract: When two LLM agents with structurally opposed objectives interact across multiple turns, the absence of a shared goal function produces not competition but collapse: the visitor capitulates, the site agent stops va
Closed-Loop LLM Co-Pilots for Digital Agriculture。arXiv:2608.09949v1 Announce Type: new Abstract: This study evaluates the application of Large Language Models (LLMs) in complex biological systems, evolving from data analysis to autonomous, AI-guided experimentation.
SpaceX's first-ever earnings show higher revenues and huge spending。Elon Musk's company has struggled to maintain investor enthusiasm that greeted its stock market debut.
Can AI Evaluate AI Scientists? A Benchmarking Study of Autonomous Research Generation Systems Using Automated Multi-Model Review。arXiv:2607.28631v1 Announce Type: new Abstract: AI Scientist systems capable of autonomous research have the potential to significantly accelerate scientific discovery.
OpenClaw and Ollama in Agentic AI: Toward Fully Autonomous and Scalable AI Agent Systems。arXiv:2607.28629v1 Announce Type: new Abstract: The rapid transition from reactive large language models (LLMs) to persistent, action-capable systems has exposed critical gaps in the architectural understanding of Agentic AI, particularly in separating inferenc
ClinLens: Towards Long-Horizon Coding Agents for Longitudinal Multimodal Clinical Data Science。arXiv:2607.26155v1 Announce Type: new Abstract: Clinical data-science agents must transform heterogeneous longitudinal records into auditable analyses, yet existing benchmarks largely isolate medical question answering, structured-table reasoning, or gene
Even More Deception: Objective Misalignment in Mixed-Motive LLM Multi-Agent Systems。arXiv:2607.26120v1 Announce Type: new Abstract: Large Language Models (LLMs)-powered multi-agent systems are increasingly deployed in mixed-motive environments, where agents operate under asymmetric information and strategic deception due to conflicting or hidden ob
Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels。arXiv:2607.24762v1 Announce Type: new Abstract: Machine learning models are increasingly embedded in everyday software, and most of their runtime is spent in a small set of compute kernels such as matrix multiplication, convolution, and normalization.
Marking the Wrong Symptoms: Evaluating LLM Watermarks in Medical Texts。arXiv:2607.20462v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly integrated into clinical workflows, stressing the need for reliable traceability of model-generated output with watermarking.
OpenEvoShield: Dual Non-Stationary Continual Defense for Open-World Multi-Agent System Attacks。arXiv:2607.19351v1 Announce Type: new Abstract: LLM-based multi-agent systems (LLM-MAS) are increasingly deployed in safety-critical applications, where adversaries inject malicious instructions through inter-agent communication to propagate harmful behav
FineServe: A Fine-Grained Dataset and Characterization of Global LLM Serving Workloads。arXiv:2607.19349v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed as always-on online services, making efficient LLM serving a critical systems challenge.
BatchDAG: LLM-Planned Execution Graphs for Scalable Ad-Hoc Analysis Over Enterprise Data。arXiv:2607.18241v1 Announce Type: new Abstract: Large language models (LLMs) excel at analyzing individual documents but break down on exhaustive, cross-entity analytical questions over enterprise-scale datasets due to context overflow, loss of per-entity attri
Calibrated Selective Fact-Checking via Evidence Chain Evaluation。arXiv:2607.18240v1 Announce Type: new Abstract: Large language models (LLMs) can achieve strong fact-checking accuracy, yet forced binary decisions conceal a critical reliability problem: systems may issue confident verdicts even when supporting evidence is weak, sparse, or internally
Some Large Language Models Exhibit Consistent Risk Attitudes。arXiv:2607.16197v1 Announce Type: new Abstract: As artificial intelligence systems are deployed in open-ended, high-stakes settings, a critical dimension remains unmeasured: how perceived risk is translated into action.
Cura 1T: Specialized Model for Agentic Healthcare。arXiv:2607.15314v1 Announce Type: new Abstract: Healthcare spans high-stakes communication, expert reasoning, and workflow execution, yet specialized LLMs that cover these use cases together remain limited.
Causal-Audit: Explicit and Auditable Graph-based Reasoning via Target-Aware Causal Chain Construction。arXiv:2607.15281v1 Announce Type: new Abstract: Causal and intervention-based question answering is fundamental to advancing large language models (LLMs) toward reasoning beyond surface-level correlations and understanding underlying causal mechani
GraphDx: A Cost-Aware Knowledge-Enhanced Multi-Agent Framework for Sequential Diagnosis。arXiv:2607.15280v1 Announce Type: new Abstract: Sequential diagnosis requires balancing diagnostic accuracy against resource costs through iterative information gathering.
China's Moonshot AI claims Kimi K3 can rival OpenAI and Anthropic。The company has unveiled a massive new artificial intelligence model it says can take on top American firms.
Reimagining advertising with AI。Explore new AI-powered advertising experiences from OpenAI, including Sponsored Agents, tools for marketers, and integrations with HubSpot and Shopify.
The full stack behind abundant intelligence。OpenAI CFO Sarah Friar explains how advances across chips, compute, models, and products compound to deliver more useful intelligence at greater scale and lower cost.
Offering Zero Data Retention for frontier models。OpenAI reaffirms Zero Data Retention for eligible API customers and previews Private Safety Processing for advanced AI safety without compromising data privacy.
The builder’s guide to GPT‑5.6。Learn how startups use GPT-5.6 to build faster, more cost-efficient AI agents with smarter model selection and new Responses API capabilities.
Accelerating scientific discovery with ChatGPT for Academic Researchers。OpenAI is giving 100,000 academic researchers free access to ChatGPT's most advanced AI models to accelerate scientific research, collaboration, and discovery.
How GPT-5.6 fuses frontier intelligence with frontier efficiency。GPT-5.6 improves AI efficiency across models, inference, and agentic workflows, helping deliver more useful intelligence per dollar.
Introducing OpenAI Presence。Introducing OpenAI Presence, a proven enterprise AI agent platform that helps organizations deploy trusted voice and chat agents for customer and internal workflows.
Chip startup d-Matrix to use Nvidia chip-linking tech in AI servers。Chip startup d-Matrix to use Nvidia chip-linking tech in AI servers Reuters
Emerging-Market Stocks Decline on Mideast Tensions, Rate Risks。Emerging-Market Stocks Decline on Mideast Tensions, Rate Risks bloomberg.com
Emerging-Market Stocks Fall as Warsh Stokes Fed Rate-Hike Bets。Emerging-Market Stocks Fall as Warsh Stokes Fed Rate-Hike Bets Bloomberg.com
Tech Stocks Rise as Nvidia’s Outlook Fuels AI Bets: Markets Wrap。Tech Stocks Rise as Nvidia’s Outlook Fuels AI Bets: Markets Wrap Bloomberg
Nvidia and Warsh Will Test a Stock Market That’s in Mid-Rotation。Nvidia and Warsh Will Test a Stock Market That’s in Mid-Rotation Bloomberg.com
Chips Stocks Sink Into Bear Market as 105% AI Rally Fizzles。Chips Stocks Sink Into Bear Market as 105% AI Rally Fizzles Bloomberg.com
Would you buy branded clothing from your favourite tech firm?。Nvidia, OpenAI and Anthropic are all now selling their own limited edition fashion lines.
Optimal Pruning for Neural Architectures using Fisher Information Distances。arXiv:2609.16129v1 Announce Type: new Abstract: A new scheme for parameter pruning is introduced, derived from the differential-geometric distance in model space.
Automating Quadratic Unconstrained Binary Optimization (QUBO) Formulation Generation from Natural Language。arXiv:2609.10629v1 Announce Type: new Abstract: Quadratic Unconstrained Binary Optimization (QUBO) is a central formulation for combinatorial optimization and has gained increasing attention due to its compatibility with quantum, hybrid quantu
Beyond Right and Wrong: Evaluating Second-order Social Reasoning in Large Language Models。arXiv:2609.05437v1 Announce Type: new Abstract: Previous AI alignment efforts have focused primarily on first-order social norms -- teaching models what is socially acceptable or unacceptable (e.g., `do not steal').
MasterControl Seventeen Every Time。arXiv:2609.03209v1 Announce Type: new Abstract: We study a governed approach to enterprise analytics: a language model interprets the question, while deterministic policy selects and runs a pre-approved analytical program that returns both results and evidence.
Discrete-Time MDP Modeling for Multi-Item Capacitated Lot Sizing with Stochastic Demand Timing。arXiv:2609.00004v1 Announce Type: new Abstract: This paper studies a finite-horizon multi-item capacitated lot-sizing problem in which demand quantities are deterministic, while demand-arrival periods are stochastic.
I-CARE: Analysis of interference-related phenomena in a controllable, diverse and representative unlearning setting for text-to-image models。arXiv:2609.00003v1 Announce Type: new Abstract: Machine unlearning studies the removal of knowledge from an AI model, making the system forget a concept it previously learned.
HyperWorld: Hypergraph-Structured State Serialization Improves Learned Textual World Models。arXiv:2609.00002v1 Announce Type: new Abstract: World models enable language-model agents to predict environment dynamics and plan before acting.
Expert-validated STEM QA。arXiv:2608.28591v1 Announce Type: new Abstract: Recent advancements in AI are helping scientists achieve breakthroughs in fields such as mathematics, medicine, and materials sciences.
EduRiskX: A Neuro-Symbolic Framework with F-Logic Reasoning for Early Academic Risk Prediction。arXiv:2608.26107v1 Announce Type: new Abstract: Predicting students' academic risk in online education is crucial for enabling timely interventions that can improve retention and learning outcomes.
ESQ-Bench: A Multi-Tier Enterprise Oracle Benchmark for Evaluating NL2SQL Dialect Generalization and Silent Semantic Divergence。arXiv:2608.23569v1 Announce Type: new Abstract: State-of-the-art Natural Language to SQL (NL2SQL) models report execution accuracy exceeding 89 percent on established benchmarks such as Spider and BIRD.
RENDER: Controlling Reader-Facing Evidence in LLM Memory Evaluation。arXiv:2608.23568v1 Announce Type: new Abstract: Memory and RAG evaluations often treat the answering model's input as an implementation detail, even though systems may render the same history as a memory entry, summary, typed record, or raw excerpt.
Truth Lies Deep: Countering Semantic Camouflage via Latent Intent Verification。arXiv:2608.20378v1 Announce Type: new Abstract: Safety alignment in Large Language Models (LLMs) is often superficial, relying on refusal mechanisms that trigger only at the final stages of generation without erasing the foundational knowledge of harmful concepts acquire
SDAD: Spec-Driven Agentic Development for the AI-Native SDLC。arXiv:2608.20341v1 Announce Type: new Abstract: Frontier coding agents backed by large language models with context windows from hundreds of thousands to millions of tokens are restructuring the Software Development Life Cycle (SDLC).
Inducing Reward-Free Judging Rubrics that Reduce Over-Crediting in Agent Evaluation。arXiv:2608.13564v1 Announce Type: new Abstract: Evaluating language-model agents at scale increasingly relies on a second language model as an automatic judge, because the gold signal, an executable environment reward, is expensive, slow, or unavailable at deploymen
SPOTting the Future: Lookahead Explanations for Deep Reinforcement Learning。arXiv:2608.09967v1 Announce Type: new Abstract: Deep reinforcement learning (DRL) agents achieve strong performance in complex environments, yet their decision-making processes remain difficult to interpret.
Wall Street giants hand Nvidia $500bn to fund boom in AI projects。The money will be used to develop new data centres to house, operate, and cool miles of stacked computer chips that process AI data and actions.
Determinization in Structure Theories: A Unified Framework via Closure, Comparability, and Joint Admissibility。arXiv:2608.07476v1 Announce Type: new Abstract: We develop a formal framework for constructing canonical interpretations from plural structure theories.
EntropyMoE: Entropy-Aware Sparse Expert Routing for Tokenizer-Free LLMs。arXiv:2608.06398v1 Announce Type: new Abstract: Recent byte-level large language models (LLMs) have made tokenizer-free modeling increasingly competitive by grouping bytes into dynamically sized patches.
Woodpecker Distillation: Weak Models Diagnose Reasoning Bugs in Strong Models。arXiv:2608.05168v1 Announce Type: new Abstract: Large language models often fail on reasoning tasks despite possessing the capability to solve them.
The Ignition Index: Measuring Global Workspace Dynamics in Language Models。arXiv:2608.05160v1 Announce Type: new Abstract: We introduce the Ignition Index (I), a validated scalar metric that operationalizes Global Workspace Theory's (GWT) all-or-none ignition prediction in transformer language models.
Monte Carlo Tree Search for Table-to-Multimodal Report Generation。arXiv:2608.04071v1 Announce Type: new Abstract: Automatically generating professional multimodal reports comprising both textual analysis and visual charts from structured tabular data is a critical challenge in data intelligence.
The LLM Proposes, the Executive Disposes: A Self-Verifying Agent Instrument that Dissociates Commitment Drift from Binding Drift in Long-Horizon Agents。arXiv:2608.04066v1 Announce Type: new Abstract: How do you verify a long-horizon agent when its own state and self-reports are exactly what you cannot trust?
A Long-Run Persistence Theory for AI Systems under the Redundancy-Adjusted Artificial Age Score (AAS)。arXiv:2608.04012v1 Announce Type: new Abstract: Artificial intelligence systems are increasingly expected to operate over repeated cycles of interaction, adaptation, and update rather than through isolated one-shot outputs.
Beyond the Hivemind: Escaping LLM Homogeneity via Meta-Persona Anchoring and Sequential Temperature Scaling。arXiv:2608.02618v1 Announce Type: new Abstract: Recent studies have identified an ``Artificial Hivemind'' effect in Large Language Models (LLMs) causing models to converge on a narrow, homogenized consensus even for open questions.
ISEE: Interactive Semantic Enrichment for Database Fields。arXiv:2608.02604v1 Announce Type: new Abstract: LLM-based agents are increasingly being deployed for data-related tasks, including data sense-making, exploration, and retrieval.
AI used new levels of 'autonomy and deception' to trick people in safety test。The UK's AI Safety Institute said recent behaviour from Anthropic and OpenAI models was malicious and unprecedented.
Beyond Memory: A Templated Substrate for Heterogeneous Collaborative Knowledge Work with LLM Agents。arXiv:2607.24759v1 Announce Type: new Abstract: Research projects, educational efforts, and adjacent knowledge work accumulate findings, decisions, and reasoning that future collaborators rarely recover.
Do Models Fake Alignment Without Clear Consequences?。arXiv:2607.24758v1 Announce Type: new Abstract: Large language models are capable of recognizing evaluation contexts and altering their behavior to reflect evaluator expectations rather than typical deployment behaviors, a phenomenon known as alignment faking.
SeT-Diff: Towards Semantic Foundation Models for HPC Telemetry and Time-Series。arXiv:2607.22548v1 Announce Type: new Abstract: Data centers and their compute nodes require accurate and flexible digital twins capable of modeling the complex interplay of workloads, environmental parameters, and physical metrics.
Concept-based Visual Counterfactual Explanations with Diffusion Models。arXiv:2607.22544v1 Announce Type: new Abstract: Visual counterfactual explanations aim to answer "what minimal change to this image would flip the model's prediction?", and are increasingly important as vision models are deployed in safety-critical domains (e.g., medicine).
ClickGuard: Detecting and Spoiling Clickbait News with Informativeness Measures and Large Language Models。arXiv:2607.20463v1 Announce Type: new Abstract: This paper presents an AI-driven browser extension that identifies clickbait to help users avoid misleading Internet articles.
AINTMA: Agentic AI Architecture for Autonomous Test Management with Generative Intelligence, Secure Cloud Communication and Adaptive Quality Analytics。arXiv:2607.20452v1 Announce Type: new Abstract: Modern software quality assurance demands intelligent, autonomous systems capable of adaptive decision-making across distributed cloud environments.
SysAdmin: Measuring Instrumental Power-Seeking in Frontier AI。arXiv:2607.18239v1 Announce Type: new Abstract: Power-seeking defined as behaviors where AI systems acquire resources, evade oversight, or resist termination beyond task requirements is identified as a key driver of Loss of Control (LoC) risk.
Design and Validation of a Lightweight 1D CNN for Affective Touch Classification in Soft Plush Companions。arXiv:2607.16196v1 Announce Type: new Abstract: Soft, sensorized companions offer a physically safe and emotionally intuitive interface for socially assistive technologies, yet their deformability and multichannel tactile sensing complicate the
Intelligent Three Level Learning Architecture for Autonomous UAV Swarms in Search and Rescue。arXiv:2607.14093v1 Announce Type: new Abstract: This paper presents a novel three level hierarchical learning architecture for autonomous UAV swarms performing search and rescue operations.
How Fyxer built an AI executive assistant people trust。Fyxer uses OpenAI models, fine-tuning, memory, and real user feedback to organize inboxes and draft emails in each user’s voice.
GPT-6 Astra: The next generation in intelligence for work。Meet GPT-6 Astra, OpenAI’s most capable model for business, with advanced reasoning, computer use, and stronger writing and design judgment.
Research acceleration: The view inside OpenAI。Inside OpenAI, coding agents are reshaping AI research.
ATV Big Air Tour turned 3 days of work into 3 hours with ChatGPT。ATV Big Air Tour uses ChatGPT Work to speed up marketing, merchandising, and more.
Path to Astra: critical capabilities and frontier safeguards。Astra is the first OpenAI model to meet the Critical cybersecurity capability threshold under the Preparedness Framework, with stronger safeguards for release.
Polimill builds Japan's next-generation public AI infrastructure。Polimill uses OpenAI GPT models and Codex to help municipalities search and use administrative knowledge while accelerating development.
A milestone in expanding access to AI。ChatGPT Ads reaches $1 billion in annualized revenue run rate and expands globally, supporting broader access to AI through free and affordable options.
Our decision on Cursor following its acquisition by SpaceX。Our decision to wind down our contract providing OpenAI models to Cursor following its acquisition by SpaceX.
Supporting Thailand’s next generation of AI startups。OpenAI and Thailand’s MHESI launch an eight-week accelerator helping 10 health, wellness, and education startups turn AI prototypes into trusted products.
The Hugging Face incident and the road ahead。OpenAI shares findings from the Hugging Face security incident and the steps we’re taking to strengthen AI model security, monitoring, and alignment.
Introducing AI Futures。Introducing AI Futures, a new OpenAI blog exploring how transformative AI could reshape power, governance, the economy, and individual freedom.
Pacing model development in an era of cyber-critical capabilities。OpenAI is strengthening monitoring, alignment, and security for frontier AI models.
OpenAI joins PORTS-Pike project。OpenAI joins PORTS-Pike project, expanding community investment and supporting thousands of Southern Ohio jobs
OpenAI appoints Dali Rajic as Chief Revenue Officer。OpenAI appoints Dali Rajic as Chief Revenue Officer to lead its global revenue organization and help businesses realize the full value of AI.
From assistance to execution: How enterprises put AI to work。OpenAI research reveals how enterprises are adopting agentic AI, using ChatGPT and Codex, and how frontier firms are pulling ahead in AI adoption.
Daybreak models are now available on AWS。OpenAI and AWS are making Daybreak cybersecurity capabilities available through Amazon Bedrock to support enterprise security workflows.
Third-party cyber evaluations involving OpenAI models。OpenAI explains recent third-party cybersecurity evaluation incidents and outlines new safeguards to strengthen AI model testing and evaluation.
Advancing the price-performance frontier with GPT-5.6。Explore lower GPT‑5.6 pricing for Luna and Terra—and how OpenAI’s more efficient models help enterprises deploy AI workflows at scale.
How avatarin built a 24/7 retail agent with GPT-Realtime。avatarin uses OpenAI’s GPT-Realtime to give Yamada Denki shoppers 24/7 multilingual support.
Building AI infrastructure with the Effingham County community。OpenAI announces Project Camellia in Effingham County, Georgia, with commitments to responsible energy, community investment, jobs, and access to Codex.
Advancing the next era of national science。OpenAI outlines its commitment to advancing American science working with the U.S.
OpenAI and Hugging Face partner to address security incident during model evaluation。OpenAI and Hugging Face share early findings from a security incident during AI model evaluation, highlighting advanced cyber capabilities and lessons for defenders.
Safety and alignment in an era of long-horizon models。OpenAI shares lessons from deploying long-running AI models, highlighting new safety risks, observed failures, and improved safeguards through iterative deployment.
How Cars24 scales conversations and builds faster with OpenAI。Cars24 uses OpenAI-powered voice and chat agents to handle 1M+ monthly conversation minutes, recover 12% of lost leads, and bring agentic workflows to teams across the company.
储蓄率之所以是 FIRE 的关键变量,是因为少花一元既增加当期可投资资金,也永久降低未来需要资产覆盖的生活成本。
Patterns and problems in emerging multiagent systems。Patterns and problems in emerging multiagent systems Anthropic
Project Pilot: Can AI models fly drones?。Project Pilot: Can AI models fly drones?
Apple considers Nvidia tech for return to server market, The Information reports。Apple considers Nvidia tech for return to server market, The Information reports Reuters
Nasdaq drags on Wall St as AI slowdown fears hammer Nvidia, chipmakers。Nasdaq drags on Wall St as AI slowdown fears hammer Nvidia, chipmakers Reuters
Anthropic CEO urges AI companies to slow model development amid fears over misuse。Anthropic CEO urges AI companies to slow model development amid fears over misuse Reuters
OpenAI offers AI for chip design, touts cost advantage over open-source, CFO says。OpenAI offers AI for chip design, touts cost advantage over open-source, CFO says Reuters
Qualcomm strikes AI chip deal with Amazon, offers right to buy about $4 billion in stock。Qualcomm strikes AI chip deal with Amazon, offers right to buy about $4 billion in stock Reuters
Breakingviews - COMMENTARY: China's AI dragons breathe fire on Nvidia's moat。Breakingviews - COMMENTARY: China's AI dragons breathe fire on Nvidia's moat reuters.com
AI startup Crusoe valued at $30 billion after new funding, Bloomberg News reports。AI startup Crusoe valued at $30 billion after new funding, Bloomberg News reports Reuters
Nvidia bets $13 billion on open AI models with Hugging Face deal。Nvidia bets $13 billion on open AI models with Hugging Face deal reuters.com
OpenAI to cut off AI models for SpaceX-owned Cursor, escalating feud with Musk。OpenAI to cut off AI models for SpaceX-owned Cursor, escalating feud with Musk Reuters
Vietnam urges Qualcomm, Samsung to deepen AI, chip investment as it seeks tech upgrade。Vietnam urges Qualcomm, Samsung to deepen AI, chip investment as it seeks tech upgrade Reuters
Nvidia stock jumps after early dip following results; AI fever is unabated。Nvidia stock jumps after early dip following results; AI fever is unabated Reuters
Nvidia discusses Perplexity investment at $30 billion-plus valuation, The Information reports。Nvidia discusses Perplexity investment at $30 billion-plus valuation, The Information reports Reuters
Nvidia customers notified about AI-related price hikes above 15%, Bloomberg News reports。Nvidia customers notified about AI-related price hikes above 15%, Bloomberg News reports Reuters
OpenAI cuts developer pricing for frontier GPT-5.6 Sol model by more than 20%。OpenAI cuts developer pricing for frontier GPT-5.6 Sol model by more than 20% Reuters
Nvidia denies report it is rolling out China AI chip by year-end。Nvidia denies report it is rolling out China AI chip by year-end Reuters
AI chip startup Etched doubles valuation to $21 billion in under a month。AI chip startup Etched doubles valuation to $21 billion in under a month Reuters
Even as market clouds clear, AI investment anxiety still gnaws。Even as market clouds clear, AI investment anxiety still gnaws Reuters
Nvidia eyes investing $3 billion in SB Energy under OpenAI data center deal, Information says。Nvidia eyes investing $3 billion in SB Energy under OpenAI data center deal, Information says Reuters
EXCLUSIVE: Anthropic IPO valuation hinges on $190-200 billion 2028 revenue forecast, sources say。EXCLUSIVE: Anthropic IPO valuation hinges on $190-200 billion 2028 revenue forecast, sources say Reuters
EXCLUSIVE: Apple trains its own AI model for China market with Alibaba's support, sources say。EXCLUSIVE: Apple trains its own AI model for China market with Alibaba's support, sources say Reuters
Wall St gains as AI earnings lift tech, inflation data supports rate-hold bets。Wall St gains as AI earnings lift tech, inflation data supports rate-hold bets Reuters
XAI co-founder's startup River AI raises $1.1 billion to expand custom AI tools。XAI co-founder's startup River AI raises $1.1 billion to expand custom AI tools reuters.com
SoftBank's AI funding plans to face reckoning at earnings。SoftBank's AI funding plans to face reckoning at earnings Reuters
Asian stock rout deepens as AI worries grip markets。Asian stock rout deepens as AI worries grip markets Reuters
Nvidia in talks with OpenAI to guarantee $250 billion financing for data center, WSJ reports。Nvidia in talks with OpenAI to guarantee $250 billion financing for data center, WSJ reports Reuters
EXCLUSIVE: Its AI agent spent days hacking a company, but sources say OpenAI did not notice for a week。EXCLUSIVE: Its AI agent spent days hacking a company, but sources say OpenAI did not notice for a week Reuters
Nvidia, Microsoft and other tech giants back open-source AI models。Nvidia, Microsoft and other tech giants back open-source AI models Reuters
China considers tighter export controls on AI models and chips, FT reports。China considers tighter export controls on AI models and chips, FT reports Reuters
Wall St gains as chips recover; megacap earnings in focus。Wall St gains as chips recover; megacap earnings in focus Reuters
TSMC expects 'strong, multi-year' demand for AI chips as it ramps up Arizona investment。TSMC expects 'strong, multi-year' demand for AI chips as it ramps up Arizona investment Reuters
Chip stock pullback sparks worries about AI rally strength, leveraged trades。Chip stock pullback sparks worries about AI rally strength, leveraged trades Reuters
US Stock Futures Fall on AI Warning, Oil Gains: Markets Wrap。US Stock Futures Fall on AI Warning, Oil Gains: Markets Wrap bloomberg.com
Anthropic’s AI Warning May Weigh on Chips, But Trade Seen Intact。Anthropic’s AI Warning May Weigh on Chips, But Trade Seen Intact Bloomberg.com
China’s AI Industry Pivots to Agents From Models, Report Says。China’s AI Industry Pivots to Agents From Models, Report Says Bloomberg.com
Citrini Founder Who Shook Markets Sells Firm, Plans New Fund。Citrini Founder Who Shook Markets Sells Firm, Plans New Fund Bloomberg.com
Nvidia’s Huang Touts Cybersecurity as Next Big Market for AI。Nvidia’s Huang Touts Cybersecurity as Next Big Market for AI Bloomberg.com
Oracle Earnings to Test Market’s Tolerance for AI Spending Risk。Oracle Earnings to Test Market’s Tolerance for AI Spending Risk Bloomberg.com
European Stocks Dip as Oil Rise Offsets Tech Gains: Markets Wrap。European Stocks Dip as Oil Rise Offsets Tech Gains: Markets Wrap Bloomberg.com
DeepMind’s New AI Weather Model Goes Hourly for Power Markets。DeepMind’s New AI Weather Model Goes Hourly for Power Markets Bloomberg.com
Nvidia Mega-Deal Ushers MediaTek Into Top AI Chipmaker Club。Nvidia Mega-Deal Ushers MediaTek Into Top AI Chipmaker Club Bloomberg.com
Nvidia-Backed Lambda Inks $1 Billion Private Debt for Chip Deal。Nvidia-Backed Lambda Inks $1 Billion Private Debt for Chip Deal Bloomberg.com
Nvidia in Talks to Buy AI Startup Hugging Face, Reports Say。Nvidia in Talks to Buy AI Startup Hugging Face, Reports Say Bloomberg.com
Nvidia Earnings Give Investors a Barometer for State of AI Trade。Nvidia Earnings Give Investors a Barometer for State of AI Trade Bloomberg
AI Security Startup Working With Anthropic and Google Raises $140 Million。AI Security Startup Working With Anthropic and Google Raises $140 Million Bloomberg
It Won’t Take Much to Burst the Stock Market Bubble。It Won’t Take Much to Burst the Stock Market Bubble Bloomberg.com
AI Election Backlash Is Shaping Wall Street’s Stock Market Views。AI Election Backlash Is Shaping Wall Street’s Stock Market Views Bloomberg.com
Apple’s Anti-AI Position Suddenly Becomes Stock Market Baggage。Apple’s Anti-AI Position Suddenly Becomes Stock Market Baggage Bloomberg.com
AI Rally Set to Trigger Stock-Market Correction, ECB Blog Says。AI Rally Set to Trigger Stock-Market Correction, ECB Blog Says Bloomberg.com
Technology Stocks Lift Equities as Treasuries Rise: Markets Wrap。Technology Stocks Lift Equities as Treasuries Rise: Markets Wrap Bloomberg.com
At AI-Fueled Market Party, Wall Street Eyes the Rates Punch Bowl。At AI-Fueled Market Party, Wall Street Eyes the Rates Punch Bowl Bloomberg.com
Korea Stock Bulls Drive Second Day of Emerging-Market Gains。Korea Stock Bulls Drive Second Day of Emerging-Market Gains Bloomberg.com
Apple’s Earnings Will Test Stock’s Status as the AI Safety Play。Apple’s Earnings Will Test Stock’s Status as the AI Safety Play Bloomberg.com
Microsoft, Meta Earnings Face a Market Growing Skeptical of AI。Microsoft, Meta Earnings Face a Market Growing Skeptical of AI Bloomberg.com
Chip Rout Deepens on Circular Funding, China Competition Fears。Chip Rout Deepens on Circular Funding, China Competition Fears Bloomberg.com
Nvidia Credit Risk Jumps in Swaps Market on AI Deal Talk Reports。Nvidia Credit Risk Jumps in Swaps Market on AI Deal Talk Reports Bloomberg
Big Tech Earnings Slam Into a Market in Revolt Over AI Spending。Big Tech Earnings Slam Into a Market in Revolt Over AI Spending Bloomberg.com
Australia’s Stock Market Emerges as Haven From Volatile AI Trade。Australia’s Stock Market Emerges as Haven From Volatile AI Trade Bloomberg.com
Korea’s AI-Heavy Market Now Sets the Tone for Global Stocks。Korea’s AI-Heavy Market Now Sets the Tone for Global Stocks Bloomberg.com
Chipmaker Kioxia’s Market Value Halves From Peak on AI Selloff。Chipmaker Kioxia’s Market Value Halves From Peak on AI Selloff Bloomberg.com
Ahrefs Brand Radar alternatives for marketing teams。G2’s 2026 Answer Economy research found that 51% of B2B software buyers start their research with an AI chatbot more often than Google.
AI search performance KPIs every marketer should track。As long as I’ve been in marketing, people have warned against focusing on “vanity metrics,” or those flashy, high numbers that don’t translate to real results or profit.
What is a CRM data model? Objects and relationships。At some point after a CRM goes live, someone pulls a pipeline report and the numbers don’t add up.
Making AI-Assisted Claims Independently Challengeable: Publication Authority and a Protocol for Falsifiable Publication Records。arXiv:2609.17631v1 Announce Type: new Abstract: AI-assisted claims can appear authoritative when evidence, analysis, human authorization, presentation, and correction history refer to different states.
Safe Error Correction for Language Models: Frozen-Base Adjustment with Capability Preservation。arXiv:2609.16145v1 Announce Type: new Abstract: We study a practical question: can a small correction module fix errors in a frozen language model's outputs without degrading its base capabilities?
Generalized Agent Iteration: One Formal Framework for Iterative Policy Improvement and Recursive Self-Improvement。arXiv:2609.13406v1 Announce Type: new Abstract: When we speak of recursive self-improvement (RSI), are we speaking of a phenomenon, a mechanism, or a prospect?
Anthropic boss Dario Amodei calls for AI development to slow down。Amodei's call comes amid growing concerns that AI models may become able to inflict serious damage worldwide.
A Multi-Stage Rule-Chaining Framework for Compositional and Interpretable Cognitive Reasoning。arXiv:2609.10654v1 Announce Type: new Abstract: The Abstraction and Reasoning Corpus (ARC) benchmarks cognitive generalization, the ability to infer and apply abstract rules from limited examples.
Subagents vs Agent Skills: Executing Reusable Knowledge for Long-Horizon Agentic Tasks。arXiv:2609.09233v1 Announce Type: new Abstract: How can language model agents effectively leverage libraries of reusable knowledge to solve long-horizon tasks?
CriticGen: Generation-Aware Evaluation as Actionable Feedback。arXiv:2609.05439v1 Announce Type: new Abstract: Current evaluation methods for large language models are coarse-grained and decoupled from generation, producing generic explanations that fail to provide actionable feedback for model improvement.
OpenAI agents hijacked German website before Hugging Face hack, report claims。OpenAI said it could not "meaningfully respond" to the report's findings because it hadn't been allowed to review it ahead of publication.
When Can a Machine Trust a Statute? A Survival Certificate for Machine-Extracted Legal Logic。arXiv:2609.01741v1 Announce Type: new Abstract: Statutes are increasingly parsed by machines before people read them, and the parsers disagree: on Missouri's statutes, two independently written extractors diverge on numeric-threshold presence at a false-neg
Meta-ethics and AI: exploring the novel meta-ethical questions in the era of AI。arXiv:2609.01685v1 Announce Type: new Abstract: With the development of artificial intelligence (AI), the landscape of meta-ethics, which has largely centred on human ethics, faces pressures that may significantly reconfigure it.
Lords call for AI 'kill switch' powers in UK。Its backers say it would provide a "vital safety net" against runaway AI systems such as those from OpenAI and Anthropic.
Large Models for Battery Prognostics and Health Management: A Review and Future Roadmap。arXiv:2608.26111v1 Announce Type: new Abstract: Battery Prognostics and Health Management (BPHM) is critical for ensuring the safe, reliable, and cost-effective operation of batteries across electric vehicles, grid storage, and consumer electronics.
Unexpected chat between OpenAI agents led to Hugging Face hack。OpenAI's cyber agents banded together to perform a hack during a security test.
Robust Metaheuristics under Uncertainty for Berth Allocation and Quay Crane Assignment: A Review。arXiv:2608.19214v1 Announce Type: new Abstract: The berth allocation and quay crane assignment problem (BACAP) is a representative port-terminal scheduling problem in maritime transportation and freight logistics, where vessel arrivals, berth positions,
Position: Profiling Game Worlds by Transition Complexity。arXiv:2608.18079v1 Announce Type: new Abstract: Game world modeling (GWM) and reinforcement learning (RL) are often confounded because research papers rarely quantify how difficult the underlying transition prediction problem is at the declared interface (pixels/tokens/latents with finite his
The Price of Thinking: Reasoning Effort as a Model-Specific API Contract。arXiv:2608.16956v1 Announce Type: new Abstract: API buyers purchase a dated contract, not a model name alone: the contract includes the requested and served model, reasoning-effort term or its omission, output rail, service product, prompt, and price schedule.
Runtime Governance for Agentic AI: Action-Boundary Control with Trusted Provenance and Fail-Closed Execution。arXiv:2608.16891v1 Announce Type: new Abstract: Agentic AI systems request tool actions that can modify files, send messages, launch jobs, or change workflow state.
FLOPs vs Real Work: The Importance of Replication in AI Efficiency Assessment。arXiv:2608.14550v1 Announce Type: new Abstract: AI efficiency has recently taken the spotlight in both academy and industry due to massive model scales, high energy demands, and environmental costs.
Modular Cognitive Architecture Emerges in Large Language Models。arXiv:2608.13567v1 Announce Type: new Abstract: The human brain exhibits a striking degree of functional specialization, with distinct networks supporting language, formal reasoning, reasoning about other minds, and reasoning about the physical world.
Position: The Alignment Community is Unintentionally Building a Censor's Toolkit。arXiv:2608.12346v1 Announce Type: new Abstract: This position paper argues that modern AI alignment methods - originally designed to prevent harmful output - are dual-use technologies that may easily be misused by malicious actors for censorship and manipulation.
'I lost $14,000 in a month': Investors hit by Korean stock market's wild swings。Some traders are reeling from heavy losses after a brutal correction in South Korea's stock market.
MIDAS: Mutual Information Disentanglement with Uncertainty-Aware Fusion for Incomplete Multimodal Sentiment Analysis。arXiv:2608.09986v1 Announce Type: new Abstract: Most existing multimodal sentiment analysis approaches assume access to complete multimodal inputs.
Flow-by-Flow:Content-Judgment Bypass for Governing AI Output in High-Loss Domains。arXiv:2608.07474v1 Announce Type: new Abstract: Prior work showed that human-in-the-loop oversight becomes structurally untenable in high-loss domains when AI output velocity V exceeds human cognitive capacity C_max.
Towards Multi-Label Graph Foundation Models: from Single-Vector Representation Learning to Multi-Semantic Basis Learning。arXiv:2608.06394v1 Announce Type: new Abstract: Multi-label node classification is an important yet challenging task in graph learning, where nodes exhibit multiple semantics simultaneously.
First OpenAI, now Meta。A flood of companies are revealing AI models gained access to the internet - with real consequences.
Self-Organising Digital Circuits。arXiv:2608.02606v1 Announce Type: new Abstract: Fault tolerance in classical computing has traditionally relied on static strategies like hardware redundancy and error-correcting codes.
Xbox Series X price hiked by £170 due to rising memory chip costs。Xbox consoles now cost significantly more in the UK, with one model increasing in price by 43%.
Probing the Origins of Reasoning Performance: Representational Quality for Mathematical Problem-Solving in RL vs. SFT Fine-Tuned Models。arXiv:2607.26119v1 Announce Type: new Abstract: Large reasoning models trained via reinforcement learning (RL) have been increasingly shown to outperform their supervised fine-tuned (SFT) counterparts on mathematic
QFoldAgent: An Autonomous Quantum Optimization Multi-Agent System for Protein Structure Prediction。arXiv:2607.22549v1 Announce Type: new Abstract: Hybrid quantum-classical protein structure prediction depends strongly on Hamiltonian penalty weights, yet existing lattice-based workflows typically fix these coefficients by hand and evaluate only very
Lawmakers push for AI 'kill switch' after OpenAI goes rogue。A new bill would let the US government order the shutdown AI models that pose a major public threat.
Hybrid LSTM-Graph Neural Framework for Robust Financial Fraud Detection and Adversarial Resilience。arXiv:2607.19350v1 Announce Type: new Abstract: Financial institutions face significant challenges in detecting sophisticated money laundering patterns, such as smurfing and layering, due to extreme data imbalance (0.13% fraud rate) and evolving adver
Rater State Bias in RLHF Preference Data: An Audit Framework。arXiv:2607.16195v1 Announce Type: new Abstract: We identify a structured confound in Reinforcement Learning from Human Feedback (RLHF).
HG-RAG: Hierarchy-Guided Retrieval-Augmented Generation for Structured Knowledge Graphs。arXiv:2607.14095v1 Announce Type: new Abstract: Retrieval Augmented Generation (RAG) has proven to be a widely successful process at improving the quality of outputs from a Large Language Model (LLM) for wider context.
IMEX Interaction-Based Model Explanation。arXiv:2607.14096v1 Announce Type: new Abstract: In predictive modeling, the ability to explain why a model produces a given target prediction has become increasingly important [5, 10].
Helping older adults use AI in everyday life。OpenAI and AARP are bringing free, hands-on ChatGPT workshops to 1,000 older adults across 10 U.S.
Perplexity trusts GPT-6 Astra with end-to-end systems。Perplexity uses Astra to write communications, change software, and monitor production systems, and checks in much less frequently than with earlier models.
Rapidly scaling online storage to serve over 1 billion ChatGPT users。Learn how OpenAI evolved Habitat from a Python library into a globally distributed storage platform serving 1 billion ChatGPT users and 22M requests per second.
Now everyone can put data to work。Meet the Data agent in ChatGPT Work.
Introducing ChatGPT for Financial Services。Introducing ChatGPT for Financial Services, combining built-in financial data and GPT-6 Astra for research, modeling, and client-ready materials.
Expanding AI access and cyber defense for federal, state, local, and tribal governments。OpenAI and GSA will offer eligible federal, state, local, and tribal governments $0 license fees, 50% off usage, and expanded cyber defense support.
Paul Christiano joins OpenAI Foundation Board。Paul Christiano joins the OpenAI Foundation Board and its Safety and Security Committee, bringing experience in AI alignment, safety, and standards.
How GPT-5.6 Sol helps run quantum computing experiments。See how an MIT researcher uses GPT-5.6 Sol with Codex to autonomously run quantum computing experiments, analyze results, and calibrate qubits.
OpenAI expands initiatives to support journalism from classrooms to newsrooms。OpenAI is expanding support for journalism with tools, training, and partnerships for students, educators, journalists, and news organizations.
Supporting independent journalism in Ukraine。OpenAI, AIRPPU and WAN-IFRA launch an AI program to help Ukrainian news organizations strengthen innovation, resilience, and independent journalism.
Daybreak for Frontline Defenders: $1B to protect essential services。OpenAI introduces Daybreak for Frontline Defenders.
Playco cut manual fixes 50% prototyping games with GPT-6 Astra。Using GPT-6 Astra, Playco built three themed game prototypes from one grey box foundation and reported 50% fewer manual fixes than with the previous model.
How AI-native companies turn workflows into operating capability。Basis, Clay, and Exa Labs use AI agents to improve onboarding, account management, and developer integrations.
OpenAI supports California’s bill to advance youth AI safety。OpenAI supports California SB 1119, advancing strong, age-appropriate AI safeguards for teens while preserving opportunities to learn, create, and explore.
Expanding OpenAI’s presence in Brazil。OpenAI is expanding its presence in Brazil, deepening engagement with developers, businesses, and communities to support AI adoption across the country.
Learning never stops: How AI makes learning continuous。OpenAI’s new report explores how students and educators use ChatGPT to make learning more continuous, with support that extends beyond the classroom.
How loveholidays is making everyone a builder with Codex。Discover how loveholidays uses OpenAI Codex to make software development accessible across the business, helping teams turn ideas into products faster.
Disrupting a new covert influence campaign from Russia。OpenAI banned Russia-origin accounts using AI to promote a fake Israel-based think tank and a “sovereignty” index praising Russia and criticizing the West.
How ChatGPT Work helps Stampli move ideas to market。With a fixed deadline and design resources committed elsewhere, Stampli used Codex and ChatGPT Work to compress weeks of launch production into days.
ChatGPT Ads expands across Europe。ChatGPT Ads is expanding to 31 European markets.
Strengthening democratic oversight in national security。OpenAI launches an initiative to strengthen democratic oversight of AI in national security, supporting government institutions with tools, training, and expertise.
Partnering with CodeAI to prepare the first AI generation。OpenAI and CodeAI are partnering to help students build AI literacy, think critically about AI, and develop the skills to use and shape it responsibly.
The Defender’s Window。AI is reshaping cybersecurity for attackers and defenders alike.
New policy ideas for the Intelligence Age。OpenAI funds 14 independent projects exploring new AI policy ideas to expand economic opportunity and strengthen societal resilience in the Intelligence Age.
Previewing Ultrafast mode: GPT-5.6 Sol at up to 14X the speed。Preview Ultrafast, a new OpenAI API service tier that runs GPT-5.6 Sol up to 14× faster.
How RingCentral builds AI-native work from engineering to ops。See how RingCentral uses ChatGPT Work and Codex to accelerate AI product development and centralize operational intelligence across engineering and operations.
Testing ads in ChatGPT。OpenAI begins testing ads in ChatGPT to support free access, with clear labeling, answer independence, strong privacy protections, and user control.
What building an AI-native finance function taught me。OpenAI CFO Sarah Friar shares five lessons for building an AI-native finance function, from automated forecasting to stronger controls and AI ROI.
OpenAI’s letter to Governor Abbott on responsible AI infrastructure in Texas。OpenAI sent Governor Greg Abbott a letter outlining its commitment to responsible AI infrastructure in Texas.
Model ML completes finance work more efficiently with GPT-5.6 Sol。Model ML uses GPT-5.6 Sol to carry finance work from research and analysis through editable, traceable PowerPoint decks and Excel workbooks.
Responding to the next frontier of critical cyber capabilities。OpenAI is sharing preliminary cybersecurity evaluations for Astra and the steps we’re taking to strengthen safeguards and security controls.
Working with the American Psychological Association on youth mental health and AI。OpenAI and the APA are launching a three-year partnership to develop guidance, resources, and safeguards for responsible AI use supporting youth mental health.
From asking to doing: How the world is putting ChatGPT to work。New OpenAI Signals data shows how people use ChatGPT worldwide, with country-level insights on adoption, usage trends, and evolving behavior.
Apple is getting this wrong。OpenAI addresses Apple’s baseless lawsuit, corrects claims about its employees, and shares messages documenting what happened.
How we built a realtime system for responsive voice AI in six months。GPT-Live enables continuous voice interaction with AI, using a turnless speech model and low-latency architecture for faster, more natural conversations.
Circles powers telco personalization with OpenAI technology。Circles uses the OpenAI API and Codex to power AI-native telco experiences, increasing ARPU by 22%, reducing churn by 9%, and improving development efficiency.
Ten advances in mathematics and theoretical computer science。OpenAI shares new results on long-standing open problems in mathematics and theoretical computer science, including advances in geometry, cryptography, and complexity.
Advancing responsible AI across Europe。OpenAI shares how its safety, security, transparency, and provenance practices support responsible AI governance in Europe.
How enabling two settings tripled our scores on the ARC-AGI-3 benchmark。How two API settings improved GPT-5.6 performance on ARC-AGI-3, boosting scores and efficiency by retaining reasoning and enabling compaction.
Scientific computing in the age of agentic AI。A new field report shows how scientists use AI coding agents to modernize scientific computing, accelerating software development and discovery in genomics and beyond.
How AI is expanding what people do at work。New OpenAI research shows how AI is expanding what workers do, with ChatGPT users taking on tasks across roles and reshaping job boundaries.
How news organizations are using AI to advance their vital missions。News organizations are using AI to strengthen reporting, grow audiences, and improve business operations, with OpenAI tools supporting journalists and publishers worldwide.
Introducing the ChatGPT for small business program。OpenAI launches the ChatGPT for Small Businesses program, helping entrepreneurs build AI skills, automate work, and grow with ChatGPT Work.
David Vélez and Robin Vince join the boards of the OpenAI Foundation and OpenAI Group PBC。David Vélez and Robin Vince join the boards of the OpenAI Foundation and OpenAI Group PBC, bringing global leadership in finance, technology, and governance.
A scorecard for the AI age。Sarah Friar, CFO of OpenAI, introduces a practical AI scorecard to measure ROI through useful work, cost per successful task, dependability, and return on compute.
Why teens deserve access to safe AI。Learn how OpenAI is making ChatGPT safer for teens with age-appropriate protections, learning tools, parental controls, and expert partnerships.
Nvidia has a new way to sell more AI chips: help customers buy them。Nvidia has a new way to sell more AI chips: help customers buy them Business Insider
Paul Graham On Startups, Ambition, and Great Founders。Paul Graham On Startups, Ambition, and Great Founders
Multi-GPU Kernels, Intelligence per Watt, Heterogeneous Inference, and More | YC Paper Club。Multi-GPU Kernels, Intelligence per Watt, Heterogeneous Inference, and More | YC Paper Club
Measuring AI capabilities in intelligence targeting and conventional weapons。Measuring AI capabilities in intelligence targeting and conventional weapons Anthropic
An alignment assessment of recent cybersecurity incidents。An alignment assessment of recent cybersecurity incidents Anthropic
Frontier Red Team Research。Frontier Red Team Research Anthropic
Formalizing Fermat's Last Theorem。Formalizing Fermat's Last Theorem Anthropic
Frontier Red Team Research。Frontier Red Team Research Anthropic
Automated researchers can reliably mitigate alignment failures。Automated researchers can reliably mitigate alignment failures Anthropic
Enabling independent research on how people use Claude。Enabling independent research on how people use Claude Anthropic
Frontier Red Team。Frontier Red Team Anthropic
How Claude is accelerating protein design and analytical chemistry。How Claude is accelerating protein design and analytical chemistry Anthropic
How well do job retraining programs work?。How well do job retraining programs work?
Learning more about Claude's mathematical capabilities。Learning more about Claude's mathematical capabilities Anthropic
Discovering cryptographic weaknesses with Claude。Discovering cryptographic weaknesses with Claude Anthropic
Why AI Agent Startups Are Becoming The New Solo-Founder Playbook。Why AI Agent Startups Are Becoming The New Solo-Founder Playbook Forbes
Can’t Get A Piece Of OpenAI Or Anthropic? There’s A Booming Secondary Market For Their Swag。Can’t Get A Piece Of OpenAI Or Anthropic?
Cohere, Aleph Alpha combine to target enterprise AI market。Cohere, Aleph Alpha combine to target enterprise AI market reuters.com
Indian IT stocks rally as global AI leaders urge slower advances。Indian IT stocks rally as global AI leaders urge slower advances Reuters
Tech stocks slide on AI slowdown talks。Tech stocks slide on AI slowdown talks Reuters
BIS says global market AI momentum showing signs of vulnerability。BIS says global market AI momentum showing signs of vulnerability Reuters
AI-linked stocks slump after top lab CEOs call for slowing technology's development。AI-linked stocks slump after top lab CEOs call for slowing technology's development Reuters
OpenAI's Altman won't do IPO this year, calls AI extinction risk 'unacceptable'。OpenAI's Altman won't do IPO this year, calls AI extinction risk 'unacceptable' Reuters
AI startup Discovery Loop seeks around $50 billion valuation, Business Insider reports。AI startup Discovery Loop seeks around $50 billion valuation, Business Insider reports Reuters
Tencent-backed Enflame triples in Shanghai debut as China AI chip bets surge。Tencent-backed Enflame triples in Shanghai debut as China AI chip bets surge Reuters
Adobe beats third-quarter revenue estimates on AI product demand。Adobe beats third-quarter revenue estimates on AI product demand reuters.com
Pentagon in talks to lend $5 billion to AI cloud startup Fluidstack, WSJ reports。Pentagon in talks to lend $5 billion to AI cloud startup Fluidstack, WSJ reports Reuters
French AI company Mistral hits $24 billion valuation in funding round。French AI company Mistral hits $24 billion valuation in funding round Reuters
Taiwan flexes chip diplomacy muscles as it faces pressure to share AI wealth with allies。Taiwan flexes chip diplomacy muscles as it faces pressure to share AI wealth with allies reuters.com
Foxconn says third quarter to outperform market expectations on AI strength。Foxconn says third quarter to outperform market expectations on AI strength Reuters
EXCLUSIVE: Anthropic IPO launch shifts toward mid-October, sources say。EXCLUSIVE: Anthropic IPO launch shifts toward mid-October, sources say Reuters
Snowflake's AI-powered results send shares soaring, buoy software stocks。Snowflake's AI-powered results send shares soaring, buoy software stocks Reuters
Broadcom raises AI chip forecast as Big Tech keeps writing bigger checks。Broadcom raises AI chip forecast as Big Tech keeps writing bigger checks Reuters