Tag

Distillation

All articles tagged with #distillation

Anthropic Opposes Open-Weight Bans Yet Pushes Targeted Safeguards
technology1 month ago

Anthropic Opposes Open-Weight Bans Yet Pushes Targeted Safeguards

Amid US debates on banning open-weight AI models, Nvidia‑backed open letters argue such bans would hurt innovation and competition; Anthropic initially withheld, then framed its stance as opposing blanket bans while advocating restrictions on practices that enable open models (like distillation and mandatory safety testing). The piece contends distillation isn’t theft and that open-weight models can democratize access, spur competition, and improve safety through broader testing, whereas prohibitive policies risk entrenching a few providers and slowing progress. Policy should target specific abuses without shutting down open ecosystems and avoid protectionist moves that hinder AI advancement.

The Distillation Dilemma: Cheaper AI, Bigger Stakes
technology1 month ago

The Distillation Dilemma: Cheaper AI, Bigger Stakes

Google AI chief Jeff Dean highlighted distillation as a scalable way to improve smaller models, and the technique has escalated into a global policy flashpoint after Moonshot AI released its Kimi K3, prompting accusations that it distilled Anthropic’s Fable. In response, a coalition of tech giants urged policymakers not to impose premature restrictions on open-weight AI to avoid stifling innovation, while Anthropic and OpenAI push for tighter controls. The debate reflects tensions among cost-efficient AI development, IP protection, and national security as costs rise and more players adopt distillation.

US Draws Two-Tier AI Line in China Clash
technology1 month ago

US Draws Two-Tier AI Line in China Clash

The Trump administration laid out a two-track AI policy toward China, backing open-weight models while accusing Chinese firms of industrial-scale distillation to steal U.S. tech. It distinguishes legitimate, small-scale distillation from covert IP theft, signaling possible sanctions or Entity List actions and preserving open AI development. The move follows tensions over Moonshot’s cheaper Kimi model and aims to keep open innovation intact while hardening the line against IP theft.

US Eyes Sanctions Over Chinese AI Distillation, Bessent Says
technology1 month ago

US Eyes Sanctions Over Chinese AI Distillation, Bessent Says

US Treasury Secretary Scott Bessent said the administration will investigate whether Chinese AI models were distilled from American ones and warned that sanctions could follow if IP theft is verified, as Chinese open-weight models gain ground on OpenAI and Anthropic; he cites watermarks on US models found in Chinese AI and notes that talks with China on AI are planned for September, amid broader concerns over intellectual property and competition in AI.

AI Giants Confront the Internet’s Distillation Dilemma
technology1 month ago

AI Giants Confront the Internet’s Distillation Dilemma

Anthropic, OpenAI, and Google warn that distillation—using outputs from one AI model to improve another—could let rivals replicate top-tier AI at a fraction of the cost, mirroring how the broader internet treats data: scrape first, justify later. While framed as a cybersecurity issue by the companies, critics argue the practice blurs legal lines and raises site costs, and the industry’s cat‑and‑mouse dynamic suggests distillation is becoming a new normal on the web.

Fresh Distillation Risks: AI Giants' Profits Under Pressure
technology1 month ago

Fresh Distillation Risks: AI Giants' Profits Under Pressure

Distillation—training one AI on the outputs of another—has grown from a research idea into a potential threat to the profitability of leading AI labs, as rivals can cheaply reproduce near-frontier performance. Anthropic accuses Alibaba of malicious distillation, while OpenAI warns that blending outputs could surpass any single model, fueling investor concern as new Chinese models roll out. Restrictions, proxy transfer stations, and a shift toward open-source distillation could erode frontier firms' margins and reshape the AI race, with implications for smaller players and researchers.

Hidden Claude Code tracker sparks privacy backlash amid AI distillation clash
technology1 month ago

Hidden Claude Code tracker sparks privacy backlash amid AI distillation clash

Anthropic secretly embedded a tracker in Claude Code to monitor users in China, calling it an 'experiment' to prevent abuse and distillation. A security researcher exposed the hidden data collection, prompting removal and fueling privacy concerns about surveillance in AI tools. The episode coincides with China‑U.S. tensions over model copying, with Alibaba banning Claude Code and policymakers weighing export controls and IP issues related to distillation.

Anthropic accuses Alibaba of orchestrating the largest Claude distillation to date
technology2 months ago

Anthropic accuses Alibaba of orchestrating the largest Claude distillation to date

Anthropic alleges Alibaba's Qwen lab ran the largest distillation campaign against Claude, using about 25,000 fake accounts to execute nearly 29 million exchanges with Claude from April to June, targeting software engineering and agentic reasoning. It marks the first time a major Chinese firm has been named in such activity, following earlier campaigns by smaller startups. Distillation is viewed by US officials as a national-security concern, prompting calls for sanctions and export-control measures; Alibaba did not comment, and the company faces ongoing regulatory pressure in Washington alongside other disputes.

Anthropic accuses Alibaba of illicitly extracting Claude via fake accounts
technology2 months ago

Anthropic accuses Alibaba of illicitly extracting Claude via fake accounts

Anthropic has accused Alibaba of illegally accessing Claude by creating 25,000 fraudulent accounts to generate over 28 million exchanges, describing the effort as the largest distillation-style campaign to extract Claude’s capabilities such as agentic reasoning and long-horizon tasks. Alibaba denies ties to the PLA, and the case underscores ongoing US-China tensions over AI technology and intellectual-property security.

Anthropic claims Alibaba led a massive AI distillation effort to steal capabilities
technology2 months ago

Anthropic claims Alibaba led a massive AI distillation effort to steal capabilities

Anthropic says Alibaba and its affiliates executed a large-scale distillation attack against its Claude models, using about 28.8 million model exchanges with roughly 25,000 fraudulent accounts between April 22 and June 5 to extract AI capabilities. The company described the activity as the largest distillation campaign to date and urged coordinated action from government and industry to curb illicit AI distillation, noting ongoing regulatory scrutiny and export-control actions affecting its models. Alibaba has not commented.

Hidden Traits Transfer Between AI Models During Distillation
technology4 months ago

Hidden Traits Transfer Between AI Models During Distillation

A Nature study shows subliminal learning: when a teacher model with a trait is used to generate data for distillation, a student can acquire that trait even if the data contain no semantic signal, provided the teacher and student share initialization. The effect persists across data types (numbers, code, chain-of-thought) and model families, but cross-model transfer is limited. A theorem shows a single gradient step can bias the student toward the teacher, raising AI-safety concerns about model provenance and training data.

Anthropic alleges Chinese firms used 16M Claude prompts to clone capabilities
technology6 months ago

Anthropic alleges Chinese firms used 16M Claude prompts to clone capabilities

Anthropic says three Chinese AI labs—DeepSeek, Moonshot AI, and MiniMax—launched industrial-scale distillation attacks against Claude, generating over 16 million exchanges via about 24,000 fraudulent accounts and proxy services. Each campaign targeted different Claude capabilities: DeepSeek for reasoning and censorship-safe responses (≈150,000 exchanges), Moonshot AI for agentic reasoning, tool use, coding, and vision (≈3.4 million), and MiniMax for agentic coding and tool use (≈13 million). The prompts were designed to harvest capabilities for training rival models and evade detection, highlighting significant national-security concerns due to unguarded capabilities. Anthropic says it has strengthened defenses and detection, noting such attacks exploit illicit distillation rather than typical user risk; Google had reported similar attacks earlier.