News
AI SHIFT:

China’s AI Labs Turn Self-Improving Models Into A Chip-Efficiency Test

Newsroom brief

Chinese AI teams are tying recursive self-improvement claims to research automation and kernel optimisation, but the strongest evidence still sits in narrow engineering tasks rather than full autonomous AI research.

Verified against source materialEdited by SendTech Times AI & Enterprise Desk
China’s AI Labs Turn Self-Improving Models Into A Chip-Efficiency Test
Image source: South China Morning Post

China’s AI Labs Move From Chatbots To Research Automation

Chinese AI developers are pushing recursive self-improvement from a speculative safety debate into concrete engineering claims.

The race is centred on systems that can help improve AI research itself, including model work, coding and chip-software optimisation.

The most important China signal is not a single product launch.

It is the clustering of claims from Xiaomi-linked model work, MiniMax, Alibaba's Qwen team, ByteDance and Tsinghua University researchers around AI systems that automate parts of development.

Luo Fuli leads Xiaomi's MiMo AI model work and told the Zhongguancun Forum in March that self-evolution was becoming a near-term AI priority, revising her earlier expectation from three to five years down to one to two years.

That shifts the competitive frame.

US companies such as Anthropic are still treated as coding and research-automation leaders, but Chinese teams are using the same frontier to offset hardware limits, especially where software efficiency can improve the use of scarce AI chips.

Kernel Optimisation Becomes The Hard Test

The clearest technical evidence sits in kernel work, not general claims about autonomous intelligence.

ByteDance and Tsinghua University researchers described a February system that used an AI agent to optimise kernels for Nvidia's CUDA ecosystem.

Their benchmark put the resulting workflow at 100 per cent faster than existing automation methods.

MiniMax made a more product-facing claim around its M3 model.

The company said the model optimised a production-grade FP8 GEMM kernel on Nvidia GPUs in around 24 hours without information that would let it copy an existing answer.

The same task would otherwise have required a human team up to two weeks, according to the company's description.

Alibaba's Qwen team pointed to a similar direction on its own hardware stack.

For Alibaba, the Qwen3.7-Max claim was tied to the company's in-house parallel processing unit platform: the model completed a kernel optimisation run in around 35 hours, or 10 times faster than the alternative process.

Those examples matter because kernel performance affects the cost and speed of large-model inference, where China faces continuing pressure from chip restrictions.

The Measurement Gap Remains Large

The caution is that task automation is not the same as proven recursive self-improvement.

Xu Weixian, an undergraduate AI researcher at Shanghai Jiao Tong University, warned that AI may be accelerating idea discovery without proving that the ideas are meaningful solutions.

Anthropic has made a similar distinction, noting a large gap between autonomous AI research tasks and autonomous goal-setting for research.

The market also lacks stable metrics.

London-based governance researcher Alan Chan said the field remains early and does not yet have concrete ways to measure progress.

Code-generation figures are useful but incomplete: Anthropic put Claude's contribution above 80 per cent of merged code as of May, and a 130-employee poll estimated that Mythos delivered a fourfold productivity gain.

The next watchpoint is whether Chinese and US AI labs publish stronger evidence that automation is improving research outcomes, not just producing faster code.

Until then, recursive self-improvement remains a strategic target with real engineering milestones and unresolved proof standards.

Share this article
inXf

Related articles

More
China’s AI Boom Faces A Demand Test As Chip Bottlenecks Raise The Cost Of Growth
AI

China’s AI Boom Faces A Demand Test As Chip Bottlenecks Raise The Cost Of Growth

China’s AI investment exceeds one trillion yuan a year, but Nomura economist Lu Ting says chip constraints, capital-market gaps and city-level concentration limit its wider economic lift.

Unconventional AI Tests Oscillator Models Before Power-Efficient Chip Proof
Chips & Semiconductors

Unconventional AI Tests Oscillator Models Before Power-Efficient Chip Proof

Unconventional AI has released the Un-0 model series to test oscillator-based image generation, but the work still runs on simulated oscillators rather than a physical AI accelerator.

NVIDIA’s NAIRR Support Shows Public AI Research Needs Dedicated Compute
AI

NVIDIA’s NAIRR Support Shows Public AI Research Needs Dedicated Compute

NVIDIA says the NAIRR pilot has supported more than 700 projects, showing how public research access to DGX infrastructure is becoming part of the AI compute debate.

Apple AI Architecture Puts Google And Nvidia Inside Its Privacy Test
AI

Apple AI Architecture Puts Google And Nvidia Inside Its Privacy Test

Apple is using Google and Nvidia to support its most advanced cloud AI model while trying to keep Apple Intelligence centered on private orchestration, proprietary models and on-device context.

Anthropic’s Conway Points Claude Toward Always-On AI Agents
AI

Anthropic’s Conway Points Claude Toward Always-On AI Agents

Anthropic is preparing a Claude expansion that includes Conway, Orbit, Operon, memory upgrades and multilingual voice mode. The move signals a shift from chat-based AI toward persistent assistants that can connect with external services and manage workspaces. Enterprises, developers and research teams could be affected if Claude becomes a broader agent platform.

Huawei's Tau scaling puts architecture at the center of China's AI-chip push
Chips & Semiconductors

Huawei's Tau scaling puts architecture at the center of China's AI-chip push

Huawei proposed Tau scaling, a framework focused on shorter signal paths rather than only smaller transistors. LogicFolding is planned for Kirin chips in fall and winter 2026, with Huawei claiming a density jump to 238 million transistors per square millimeter. The test is whether architecture-led gains can be validated in commercial chips amid export-control limits on advanced tools.

Nota Runs VLA Robotics Model in Real Time on Qualcomm Edge AI Hardware
AI

Nota Runs VLA Robotics Model in Real Time on Qualcomm Edge AI Hardware

Nota demonstrated real-time operation of a vision-language-action robotics model on Qualcomm Dragonwing edge AI hardware. The company reduced the model action-head processing time from 218 milliseconds to 31 milliseconds while keeping task success nearly unchanged. The demo points to a path for physical AI systems that can run closer to robots rather than relying mainly on GPU servers or cloud infrastructure.

Chinese Brain-Mapping Chip Claims 478 Times NVIDIA A100 Speed Without Clinical Proof
Chips & Semiconductors

Chinese Brain-Mapping Chip Claims 478 Times NVIDIA A100 Speed Without Clinical Proof

A Peking University and Chinese Academy of Sciences team reported a 40-nanometre computing-in-memory chip that maps the brain’s folded cortex in less than 0.5 seconds, while the public record still lacks clinical validation or a commercial release plan.

Keep Reading

More Stories

Latest
Ramp And FIS Add AI Controls For Spend And Payments SecurityFintech & Digital PaymentsJul 27, 2026Ramp And FIS Add AI Controls For Spend And Payments SecurityRamp launched controls for AI spending while FIS joined Anthropic Project Glasswing to test Mythos 5 against payment-infrastructure vulnerabilities, leaving adoption and measured security results outside the public record.UAE Statistics Programme Targets AI-Ready Official DataCapital & PolicyJul 27, 2026UAE Statistics Programme Targets AI-Ready Official DataThe UAE has approved a national programme to improve official statistics as government agencies seek higher-quality data for economic policy and artificial intelligence systems.SourTrade Malvertising Makes Browsers Assemble Windows MalwareCybersecurityJul 27, 2026SourTrade Malvertising Makes Browsers Assemble Windows MalwareThe Hacker News reported that Confiant analysed SourTrade, a malvertising campaign that uses fake trading pages and browser-side assembly to vary Windows malware files for retail trading and crypto targets.Lumen Tracks Proxy Botnets Approaching 60 Million IPsCybersecurityJul 27, 2026Lumen Tracks Proxy Botnets Approaching 60 Million IPsCyberScoop reported that Lumen Black Lotus Labs is tracking malicious residential proxy botnets approaching 60 million victim IP addresses, with takedowns failing to stop rapid rebuilds.Shopify Reworks Theme Code For AI Agents And ReadabilityAIJul 27, 2026Shopify Reworks Theme Code For AI Agents And ReadabilityShopify is preparing a cleaner storefront theme that shifts from JSON-heavy configuration toward mostly HTML and Liquid, as AI-assisted theme editing makes readable code and explicit contracts more important.Empery Invests $20 Million In Cardinal Data PowerCloud & Data CentersJul 27, 2026Empery Invests $20 Million In Cardinal Data PowerEmpery Digital invested $20 million in Cardinal Data Power, backing a West Texas data centre campus plan with a 750MW first phase and potential expansion beyond 5GW.AI Distillation Debate Moves From Labs To WashingtonAIJul 27, 2026AI Distillation Debate Moves From Labs To WashingtonCNBC reported that AI distillation has become a policy fight after Moonshot AI's Kimi K3 raised questions about open-weight models, proprietary model output and U.S. restrictions.Hugging Face Adds 4-Bit Nunchaku Loading To DiffusersChips & SemiconductorsJul 27, 2026Hugging Face Adds 4-Bit Nunchaku Loading To DiffusersHugging Face added Nunchaku Lite support to Diffusers, letting developers load 4-bit diffusion checkpoints with `from_pretrained()` while using Hub-delivered CUDA kernels.CFTC Warning Narrows Prediction-Market Contract FilingsFintech & Digital PaymentsJul 27, 2026CFTC Warning Narrows Prediction-Market Contract FilingsCoinDesk reported on July 26 that the CFTC warned prediction-market operators against broad template certifications, tightening the filing burden for event contracts while courts still test the regulator’s authority.B Capital Names Ex-G42 AI Chief To Doha Investment RoleAIJul 27, 2026B Capital Names Ex-G42 AI Chief To Doha Investment RoleB Capital has appointed former G42 executive Dr Andrew Jackson as General Partner and Chief AI Officer in Doha, tying the firm’s Gulf office to AI investment governance and Stargate UAE experience.Ajman AI Agent Renews Trade Licences Through AjmanOneEconomyJul 27, 2026Ajman AI Agent Renews Trade Licences Through AjmanOneMiddle East AI News reported that Ajman used agentic AI to renew a trade licence through AjmanOne, moving a government service from chatbot support toward an automated transaction flow that still depends on governed data and system integration.AMD Helios AI Racks Bring Epyc CPUs To Nvidia Compute ChallengeChips & SemiconductorsJul 26, 2026AMD Helios AI Racks Bring Epyc CPUs To Nvidia Compute ChallengeData Center Knowledge reported that AMD moved Helios into production with MI455X GPUs, Epyc processors, Pensando networking and ROCm software, while vendor performance claims still lack third-party benchmark validation.