Hasty Briefsbeta

Bilingual

Chinese military researchers tap US AI models to train defense systems

5 hours ago
  • Chinese military researchers have used outputs from U.S. AI models like OpenAI's GPT-3.5 and Anthropic's Claude to train domestic defense AI systems via a technique called model distillation.
  • A Reuters review of over 80 Chinese academic papers and patents reveals widespread use of distillation by PLA-linked institutions to develop specialized AI for surveillance, cyber warfare, and tactical decision-making.
  • Model distillation involves using outputs from a powerful AI system to train smaller, specialized models, reducing the computing requirements compared to building frontier AI from scratch.
  • The practice is disputed, with U.S. officials accusing Chinese entities of unauthorized extraction that undermines export controls and infringes IP, while China denies reliance and accuses the U.S. of 'hegemonism'.
  • One paper from PLA Unit 96941 described using GPT-3.5 to summarize sensitive military source code and training a domestic model on those summaries for use within military networks.
  • Other examples include using Claude 3 Haiku to generate training data for social media monitoring, distilling image-processing models for drone deployment, and target recognition for maritime operations.
  • China promotes model lightweighting and edge computing to enable AI on drones and satellites, but experts note distilled models have limitations and cannot fully replicate frontier AI's broad capabilities.
  • Chinese researchers are also studying distillation as a security risk, proposing defenses against 'data-free distillation' that could reverse-engineer models.