September 12, 2026
English translation of the Korean original, prepared with AI assistance. Korean original
Top headlines
- Anthropic claims Moonshot, DeepSeek secretly diverted user requests to Claude
- Background
- Claude is a conversational AI built by Anthropic. Moonshot and DeepSeek are Chinese AI start-ups that have grown their user bases quickly with low-cost models. It has long been alleged that Chinese firms route around restrictions to connect to foreign model APIs.
- Why it matters
- Anthropic has publicly accused a competitor of secretly diverting traffic to its models. This has fuelled a dispute over breaches of API terms of service and the US-China AI rivalry.
- So what
- Companies that have adopted external AI APIs are advised to set up a procedure to check that the model actually processing their requests matches the one they contracted for.
- Build more natural voice experiences with GPT-Live-1 in the API
- Background
- OpenAI has kept extending its text-centred GPT models into voice assistants. Its earlier voice feature was a half-way form of conversation that cut out or lagged when speech overlapped.
- Why it matters
- GPT-Live-1 offers natural two-way calls through an API, in which speech overlaps and people interrupt, as humans do. This changes how call centres and voice assistants are built.
- So what
- It is time for companies building voice services to consider redesigning their existing voice UX with GPT-Live-1.
- Expanding AI access and cyber defense for federal, state, local, and tribal governments
- Background
- The GSA is the procurement agency that standardises contracts when US government bodies adopt private-sector technology. OpenAI has so far expanded into the public sector mainly through the federal government.
- Why it matters
- Under this partnership, state, local and tribal governments as well as the federal government can use AI and cyber-defence services at discounted prices. That speeds up the adoption of AI in the public sector.
- So what
- Officials at local governments and public bodies should check the terms of the discounted licences available through the GSA now, and draw up their adoption plans.
DeepSeek-V4.1-Flash offers very low prices and high performance, and is reshaping the cost structure of the LLM market. It means that efficiency, along with gains in individual model performance, has become the key competitive strength.
AI is reaching more of everyday life at an accelerating pace, through voice conversation and CRM integration built on Claude. AI has entered a stage where it goes beyond generating language and is embedded deeply in real business processes.
The contest for dominance over building huge infrastructure is likely to be the main driver of the market. Note in particular Nvidia’s analysis of the dynamics of an AI ecosystem worth $11tn.
Signals 39
Rapidly scaling online storage to serve over 1 billion ChatGPT users
OpenAI has turned its Habitat library into a globally distributed storage platform that serves 1bn users and handles 22m requests per second.
Signal — Future AI competitiveness will be decided not only by model performance but by the ability to scale the back end to withstand explosive user traffic.
OpenAI Blog
Expanding AI access and cyber defense for federal, state, local, and tribal governments
OpenAI and the GSA will offer discounted licences and cyber-defence services to US government bodies at every level (federal, state, local and tribal).
Signal — The next growth engine for AI technology will be the public sector, which demands strong security and regulatory compliance as well as technical excellence.
OpenAI Blog
Build more natural voice experiences with GPT‑Live‑1 in the API
GPT-Live-1 is a voice AI service that offers natural two-way (full-duplex) voice conversation, custom voices and phone-call support through an API.
Signal — More advanced voice AI will raise the importance of voice-specific AI chipsets (ASICs) and intensify the competition in real-time processing optimisation.
OpenAI Blog
Boots on the Ground: ‘WARDOGS’ Goes All Out on GeForce NOW at Early-Access Launch
NVIDIA's GeForce NOW is expanding its service by streaming many of the latest high-end PC games, such as WARDOGS, from the cloud.
Signal — Beyond gaming, the move to the cloud will accelerate for every professional workload that needs high precision and high-definition graphics, such as medical simulation and remote manufacturing control.
NVIDIA Blog
Nvidia’s Backstop Universe – Heads I Win, Tails Who Loses?
It analyses Nvidia's monopoly position and economic influence in the coming build-out of $11tn of AI infrastructure.
Signal — Amid the vast flows of capital and allocation of resources in the AI market, it is necessary to watch closely Nvidia's financial risk (its balance sheet) and the durability of its market hegemony.
SemiAnalysis
Anthropic Says Seven China-Based AI Labs Ran Industrial-Scale Claude Distillation Attacks - The Hacker News
Anthropic claims it was the target of a large-scale attack by Chinese research institutions that sought to reverse-engineer the intellectual property of its Claude models through distillation.
Signal — As the boundaries between countries and companies erode, securing AI sovereignty over models will emerge as a core governance issue.
Foundation model capabilities & benchmarks
Blog: Accenture Launches Accenture Trusted Wealth Ops, Powered by Salesforce and Claude, to Help Wealth Advisors Deepen Client Relationships - newsroom.accenture.com
Accenture has combined Salesforce and Claude to launch a customer relationship management (CRM) operations solution for wealth advisors.
Signal — The focus of LLM adoption is moving from verifying technical performance to delivering "tailored trust solutions" that reflect industry-specific regulation and workflows.
Foundation model capabilities & benchmarks
DeepSeek's new model sets a template for powerful LLMs that run lean - The Register
DeepSeek has unveiled a new LLM that is resource-efficient and delivers strong performance.
Signal — Improving model efficiency will become the key bottleneck in deploying AI models, and commercial use cases for lightweight models are expected to increase.
Foundation model capabilities & benchmarks
DeepSeek V4.1-Flash puts pressure on AI pricing - The Rundown AI
DeepSeek has released V4.1-Flash, a new foundation model that maximises speed and efficiency while keeping performance high.
Signal — Going forward, stable and optimised "inference cost" and "operating efficiency", rather than peak model performance, will be the core competitive strengths of AI services.
Foundation model capabilities & benchmarks
Grok vs ChatGPT vs Gemini: Hallucination Rate Compared - tech-insider.org
A comparative analysis that measures the real-world hallucination rates of leading LLMs such as Grok, ChatGPT and Gemini.
Signal — Competition on model performance will now be reorganised around "how reliable is it" (verifiability), rather than "how smart is it".
Foundation model capabilities & benchmarks
DeepSeek-V4.1-Flash debuts with $0.003/1M off-peak cached-input rate and benchmarks eclipsing GPT-5.6 Sol, Claude Opus 5 - VentureBeat
DeepSeek-V4.1-Flash launched at a very low cost ($0.003/1M) and claims benchmark performance that beats top models such as GPT-5.6 Sol and Claude Opus 5.
Signal — The key factor in AI model competition is moving beyond a simple performance lead to "performance per dollar" at a sustainable operating cost.
Foundation model capabilities & benchmarks
Watch DeepSeek's New Model Rattles Chipmakers and AI Rivals - bloomberg.com
DeepSeek is drawing market attention by releasing a new high-performing LLM as open weights.
Signal — As open-weight models rapidly catch up with commercial models on performance, the next round of competition will move to efficiency and ease of deployment.
Open model & open-weight releases
Anthropic claims Moonshot, DeepSeek secretly diverted user requests to Claude - South China Morning Post
Market competition between Anthropic's Claude and DeepSeek is intensifying, and suspicions have been raised that it is gaining share by diverting traffic.
Signal — Beyond model performance and technical superiority, a "user acquisition strategy" that combines user experience (UX) design with strategic marketing will become a key competitive strength.
Open model & open-weight releases
DeepSeek Launches V4.1-Flash With Lower Memory and API Costs - TechRepublic
DeepSeek has released V4.1-Flash, an efficient version of its model with lower memory usage and API costs.
Signal — More than performance itself, "how cost-efficiently a model can be deployed" will emerge as the core value of a foundation model.
Open model & open-weight releases
DeepSeek Ships V4.1 Flash, Sunsets V4 Pro Sept. 14 [2026] - shattered.io
DeepSeek has released V4.1 Flash, a new lightweight open-weight model, and published the end-of-support schedule for its existing models.
Signal — Specialised compact models, made ultra-light and optimised for specific tasks and environments, will become the mainstream.
Open model & open-weight releases
Open-Source AI & Open Models Reading List
A comprehensive reading list that sets out the main open-source models and their industrial implications.
Signal — Beyond the race over the scale of general-purpose models, an open-model ecosystem of lightweight, specialised and fine-tuned models optimised for specific purposes will take off in earnest.
Interconnects
Subagents vs Agent Skills: Executing Reusable Knowledge for Long-Horizon Agentic Tasks
To solve long-horizon agent tasks, it proposes reusing knowledge by calling sub-agents, instead of the existing approach of skills that load context.
Signal — The focus in building agent intelligence is shifting from a race over context-window size to a race over "hierarchical architectures" and "memory management structures".
arXiv cs.AI
An Autonomous GeoAI Agent for Arctic Eco-Navigation
It proposes a multi-agent GeoAI route-finding system for planning Arctic shipping routes that weighs operational, physical, ecological and community impacts together.
Signal — AI is developing in a direction where regulatory compliance and social responsibility become core goals, beyond functional optimisation.
arXiv cs.AI
Do Agents Know When They Succeed? Calibrating Agent Confidence from Internal Representations
It presents a method for measuring the success of complex tasks in agent systems from the model's internal representations, using trace trajectory dynamics (LTD) and action representation probes (ARP).
Signal — The paradigm for evaluating LLM performance is shifting away from the output text alone toward quantifying a system's behavioural trajectory and its confidence.
arXiv cs.AI
M3-Former: Multimodal Transformer with Mixture-of-Experts for Long-Term Vessel Trajectory Prediction
M3-Former is a multimodal Transformer framework that combines an LLM with MoE to predict the long-term trajectories of ships.
Signal — A major trend is to combine the knowledge base of general-purpose LLMs with the physical constraints or time-series information of specialised domains (physics-constrained LLMs).
arXiv cs.LG
RiVaT-Fuse: Reliability-Calibrated Variational Tensor Fusion for Multimodal Prediction under Modality Uncertainty
It proposes a variational tensor fusion framework that combines heterogeneous evidence, such as images and metadata, while accounting for modality uncertainty.
Signal — AI models will evolve beyond simply making predictions, toward quantifying the reliability of each piece of data and handling "uncertainty".
arXiv cs.LG
T1: Terminal Agent Reinforcement Learning for Long-Horizon Tasks
T1 is a large MoE-based agent model trained with RL to achieve long-term goals through complex tool calls of more than 300 turns in a real shell environment.
Signal — The focus of AI research is clearly shifting from linguistic understanding to action planning and completion in an environment.
arXiv cs.LG
CMNIE: An Information Extraction Benchmark for Chinese Military News
It presents a multi-task information extraction benchmark for Chinese military news (CMNIE), annotating events, entities and relations together.
Signal — Beyond general datasets, high-quality benchmark data specialised for particular domains that demand high security and expertise is growing in importance.
arXiv cs.CL
Think Before You Link: Rarity, Reasoning, and Retrieval in Multilingual Entity Linking
A research paper that improves multilingual entity linking for rare entities, using knowledge-graph structural metrics and a VLM with reasoning capability.
Signal — To raise trust in LLMs, the benchmark for measuring performance is trending from simple accuracy to robustness of knowledge structure.
arXiv cs.CL
LLM-Anchored Paralinguistic Enrichment for Alzheimer's Disease Detection
It proposes complementing the linguistic representations of language models (LLMs) by integrating prosodic features such as pauses and elongation (LAPE).
Signal — The commercialisation of hyper-precision multimodal models that predict disease by analysing subtle changes in human biosignals will accelerate.
arXiv cs.CL
OpenAI’s feud with mathematicians is only escalating
Leaders in mathematics have published an open letter warning that AI research threatens intellectual property.
Signal — The fundamental debate will deepen over whether machines are replacing uniquely human intellectual activity, or whether it will be redefined as collaboration with humans.
TechCrunch AI
One week left to book your exhibit table at TechCrunch Disrupt 2026
A notice that the deadline for booking exhibition booths at TechCrunch Disrupt 2026 is near, urging participants to show strong interest.
Signal — This is a sign of the times: whatever the pace of AI progress, the importance of physical, offline communities and verification is being stressed again.
TechCrunch AI
Final, final, final call for TechCrunch Disrupt 2026 Side Events
A notice that the opportunity to apply to host an official side event at TechCrunch Disrupt 2026 has closed.
Signal — The latest themes of AI's "commercialisation stage" and the areas drawing large capital, as covered at major industry conferences, will be the next major trends.
TechCrunch AI
Why is TMLR so slow in recent times [D]
An account of an academic publishing process in which the final review and acceptance notice for a paper submitted to TMLR has been delayed by more than two months.
Signal — There is a growing need for outcome-based verification systems that recognise the value of research in real time and deliver results to fit career timelines.
Reddit r/MachineLearning
ACL Sustainable Reviewing Policy [D]
To address submission overload and a shortage of reviewers, the ACL conference has made reviewer participation mandatory for each submitted paper and has capped the number of submissions it accepts.
Signal — In response to the surge in academic content, the peer review process itself will turn into a more systematic and rigorous commercial and structural system.
Reddit r/MachineLearning
Anthropic: Our Thoughts Pre S-1 Filing - CreditSights
A report in which an investment firm analyses Anthropic's financial position and market value ahead of its potential listing (S-1).
Signal — As AI companies enter a mature stage, financial sustainability and institutionalisation will matter more than technical innovation.
AI capital markets (IPOs, funding, valuations)
Nscale adds former OpenAI exec Fidji Simo to its board ahead of potential IPO - TechCrunch
Nscale has brought former OpenAI executive Fidji Simo onto its board in preparation for a potential IPO.
Signal — Watch the trend in which the growth stage of AI start-ups evolves beyond simple technical implementation to structural stability and professional management for an IPO.
AI capital markets (IPOs, funding, valuations)
AI regulation calls grow in DC after researcher's extinction warning - CNBC
As academic warnings about the dangers and potential risks of AI grow, discussion of AI regulation is intensifying in Washington, D.C.
Signal — Whether national-level AI risk assessments and AI safety certification frameworks will emerge.
AI governance & regulation (government, security)
Lawmakers are pushing for AI regulation after an Anthropic researcher's extinction warning - qz.com
Following warnings from AI researchers about existential threats, legislators around the world are accelerating the introduction of AI regulation.
Signal — Going forward, developing models that "meet regulatory and safety standards", rather than asking which model is better, will be the biggest core competitive strength.
AI governance & regulation (government, security)
Can Samsung’s HBM growth hold off rising Chinese foundry rivals? - KED Global
It analyses how Samsung Electronics' dominant growth in the HBM (high-bandwidth memory) market can withstand competitive pressure from emerging Chinese foundries.
Signal — Competition in the HBM market will intensify beyond simple memory supply into a contest over memory-logic integration (such as HBM-PIM) and packaging technology.
Custom silicon & HBM
DeepSeek Cut HBM Cache Needs 75% and SSD 87.5%. What Does It Mean for Memory? - Moomoo
Optimising for DeepSeek models has cut the HBM cache and SSD capacity that large language models require to 75% and 87.5% of previous levels, respectively.
Signal — Beyond the race over model size, "optimised deployment architectures" that can run efficiently on limited resources will be the next core competitive advantage.
Custom silicon & HBM
Kepler Computing Emerges to Build HBM Alternative Using FeRAM - TechPowerUp
The emergence of Kepler Computing, a start-up developing an alternative to high-bandwidth memory (HBM) based on FeRAM (ferroelectric RAM).
Signal — To raise computing performance, memory architecture will not be unified. A diversified structure in which several types of memory are combined as needed (heterogeneous integration) will become the mainstream trend.
Custom silicon & HBM
Hudbay Minerals Inc. (HBM:CA) Presents at Jefferies Global Industrials Conference 2026 - Slideshow - Seeking Alpha
The article covers the macro investment and demand cycle for mineral resources and industrial infrastructure.
Signal — What decides the success of the AI stack will be not only software and algorithms but also stable access to physical resources, accompanied by sustainable energy.
Custom silicon & HBM
How enterprise AI cost management works - IBM
In large-enterprise settings, the key issue in AI adoption is moving beyond model development (R&D) to managing operating costs and securing efficiency.
Signal — The next AI trend will focus not on "which model is best" but on "how it can be run cost-efficiently".
AI demand, pricing & unit economics