Anthropic-ன் Claude Fable 5.1 என்பது இன்னொரு சற்றே நல்ல chat model மட்டும் அல்ல. பல மணி நேரம் நீளும் coding, research, documents, tools மற்றும் தோல்வியடைந்த experiments ஆகியவற்றைத் தொடர்ந்து செய்யும் agent-ஆக இது உருவாக்கப்பட்டுள்ளது.

இதனை வேலைக்கு விட்டு வரக்கூடிய digital colleague போல நினைக்கலாம். இது notes வைத்துக்கொள்கிறது, output-ஐச் சரிபார்க்கிறது, நீண்ட task-ன் நோக்கத்தை மறக்காமல் இருக்க முயல்கிறது.

Source note: இங்கே உள்ள numbers Anthropic announcement மற்றும் partner reports-லிருந்து எடுக்கப்பட்டவை. இது சுயாதீன VelsTech benchmark அல்ல.

சுருக்கமாக

இதன் உண்மையான பலம் speed அல்ல

Announcement-ல் உள்ள சிறந்த examples flashy demos அல்ல. Millennium-ன் தகவல்படி, Fable 5.1 மிகவும் அரிதான crash-ன் காரணத்தை கண்டுபிடித்தது. Datadog production incidents-ல் root-cause analysis-ஐக் கண்டது. Red Hat broken builds-ன் காரணத்தை சரியாகக் கண்டதாகக் கூறுகிறது.

இத்தகைய பணிகளில் முதல் plausible answer போதாது. model hypothesis உருவாக்கி, evidence சேகரித்து, contradiction-ஐக் கவனித்து, மீண்டும் முயல வேண்டும். அதனால் verification loops response speed-ஐ விட முக்கியமானவை.

Cache-read மாற்றம்

Input $10 மற்றும் output $50 per million tokens என்ற headline pricing தொடர்கிறது. ஆனால் cache reads $0.25 per million tokens ஆகக் குறைந்துள்ளன. ஒரே repository, system prompt மற்றும் research notes-ஐ மீண்டும் மீண்டும் படிக்கும் agent-க்கு இது பெரிய சேமிப்பு.

Fable மற்றும் Mythos

Fable 5.1 பொதுப் பயன்பாட்டுக்கானது. Mythos 5.1 trusted access programs வழியாக தேர்ந்தெடுக்கப்பட்ட cyber மற்றும் life-science professionals-க்கு வழங்கப்படுகிறது. இதை downloadable “uncensored” checkpoint என நினைக்கக் கூடாது. Access மற்றும் safeguards product-ன் பகுதிகள்.

Local-AI பயனர்களுக்கு

Fable 5.1-ஐ RX 6800M, RTX அல்லது Apple Silicon-ல் இயக்க download செய்ய முடியாது. Local AI privacy, offline use மற்றும் கணிக்கக்கூடிய செலவில் வலிமையானது. Frontier reasoning தேவைப்பட்டாலும் frontier hardware வாங்க விரும்பாதவர்களுக்கு Fable பொருத்தமானது.

இதன் workflow ideas local models-க்கும் பயன்படும்: durable notes, checkpoints, ஒவ்வொரு மாற்றத்திற்குப் பிறகும் tests, மற்றும் agent-இடம் evidence கேட்பது.

என் முடிவு

Fable 5.1 ஒரு புதிய chat personality-யை விட dependable agent ஆக மாற முயல்கிறது. Coding, science, finance மற்றும் incident response ஆகியவற்றில் ஒரே pattern தெரிகிறது: model thread-ஐ இழக்காமல் தன் வேலையை verify செய்கிறது.

Sources