Close Menu
  • AI
  • Content Creation
  • Tech
  • Robotics
AI-trends.todayAI-trends.today
  • AI
  • Content Creation
  • Tech
  • Robotics
Trending
  • Supersonic Labs Releases Julia 1: A 144.3M-Parameter Open Choice Mannequin That Runs on a CPU
  • Sarvam AI Releases Saaras V4: A Speech-to-Textual content Mannequin for All 22 Indian Languages and World English
  • Meta says it’s going to run advertisements for ‘Musk’ documentary in spite of everything
  • Finish-to-Finish Multimodal Information Augmentation and Adversarial Robustness Benchmark with AugLy for Pictures, Textual content, Audio, and PyTorch
  • Meta’s Muse Is Adults-Solely. Why Does It Look Like a Children’ Toy?
  • Exa Launches Agent Extremely: A Subagent Swarm Deep Analysis API Constructed for Exhaustive Checklist Constructing
  • Liquid AI Releases LFM2.5-VL-3B-DSpark: Speculative Decoding for Imaginative and prescient-Language Fashions With As much as 3.13x Sooner Decoding
  • Thieves Stole ‘Nvidia’ Trailers. They Bought 20 Tons of Sand
AI-trends.todayAI-trends.today
Home»Tech»Supersonic Labs Releases Julia 1: A 144.3M-Parameter Open Choice Mannequin That Runs on a CPU

Supersonic Labs Releases Julia 1: A 144.3M-Parameter Open Choice Mannequin That Runs on a CPU

Tech By Gavin Wallace27/09/20263 Mins Read
Facebook Twitter LinkedIn Email
Step-by-Step Guide to Creating Synthetic Data Using the Synthetic Data
Step-by-Step Guide to Creating Synthetic Data Using the Synthetic Data
Share
Facebook Twitter LinkedIn Email

Supersonic Labs, a small AI lab from Brazil, has launched Julia 1. It’s a compact determination mannequin, not a chatbot. You cross it context, a query, and a pair of to twenty candidate solutions. It picks one and returns a chance for each choice. The mannequin has 144.3M parameters and runs on a plain CPU.

Is it deployable? Sure. The weights are on Hugging Face beneath Apache 2.0 and run domestically with Python 3.11+ on CPU or a BF16-capable GPU. An ONNX build additionally runs within the browser by way of WebGPU. A hosted API is introduced however not open but.

What Julia 1 Does

Julia 1 handles three determination varieties by means of one API:

  • alternative: choose one label from 2 to twenty described choices (classification, routing).
  • rating: return the anticipated index on an ordered rubric, reminiscent of low, medium, excessive.
  • noul: return the chance {that a} yes-or-no assertion is true.

Outcomes come again within the caller’s choice order with full softmax chances. Caller IDs reminiscent of billing are returned unchanged. The mannequin doesn’t generate textual content.

Structure and Coaching Funds

Julia 1 begins from JHU CLSP’s mmBERT-small, a 140M-parameter multilingual ModernBERT encoder educated on 1,800+ languages. Supersonic Labs saved the encoder and tokenizer, added a choice head, and educated on decision-format examples. The lab states Julia 1 is not a fine-tuned Qwen mannequin. The runtime helps 8,192 mixed tokens, however printed benchmarks used a 1,024-token restrict.

Whole cloud GPU spend for coaching and experiments was about R$540 (US$104.08). The FP32 weights occupy 550.5 MiB. The non-public coaching pipeline will not be launched. Julia 2, with the lab’s personal basis structure, is in growth.

Benchmark Outcomes

The September 24, 2026 analysis ran on H200 BF16 with strict encoding. The comparability baseline is TypeSafe’s Jev, utilizing reference values from the Jev benchmark protocol, not a brand new Jev run.

  • Typed Decisions: 73.15% (1,463/2,000) vs 72.70% reference.
  • AG Information, 4 labels: 94/100 vs 91% reference.
  • DAIR Emotion, 6 labels: 86/100 vs 48% reference.
  • Banking77, 72 labels: 64/100 vs 87% reference. That is the clear failure.
  • MASSIVE, 18 situations: 71.50% macro accuracy throughout 52 locales; 86.25% pt-PT, 86.75% en-US.

The classification pilots use solely 100 examples every. A September 25 CPU run reproduced most numbers: 72.55% on Typed Choices and 60/100 on Banking77 with 3 abstentions.

On-Machine Latency

The lab printed per-device measurements. On an Apple M4, one determination per name took a 33.15 ms median. On a Samsung SM-X510 pill by way of ONNX Runtime, the median was 203 ms with 393.1 MB peak RSS. On an Intel Core i5-1235U, AG Information selections took a 107.83 ms median. Banking77 took 3,713.54 ms as a result of it narrows 72 labels first.

On X, @supersonicai claims Julia 1 classifies 5x quicker than Jev on an i5 laptop computer. Deal with that rigorously. The Jev pilot measured Jev as a hosted service referred to as from France, so latencies are usually not like-for-like.

Interactive Explainer

‘;doc.getElementById(‘dots’).innerHTML=h;
doc.getElementById(‘rSteps’).innerHTML=”;doc.getElementById(‘rOut’).textContent=n‘+r[0]+’
‘+r[3]+’ · ‘+(d>0?’+’:”)+d+’ pp

Julia 1: ‘+fm(r[1])+’%

Jev ref: ‘+fm(r[2])+’%

ar met
Share. Facebook Twitter LinkedIn Email
Avatar
Gavin Wallace

Related Posts

Sarvam AI Releases Saaras V4: A Speech-to-Textual content Mannequin for All 22 Indian Languages and World English

26/09/2026

Finish-to-Finish Multimodal Information Augmentation and Adversarial Robustness Benchmark with AugLy for Pictures, Textual content, Audio, and PyTorch

26/09/2026

Exa Launches Agent Extremely: A Subagent Swarm Deep Analysis API Constructed for Exhaustive Checklist Constructing

26/09/2026

Liquid AI Releases LFM2.5-VL-3B-DSpark: Speculative Decoding for Imaginative and prescient-Language Fashions With As much as 3.13x Sooner Decoding

26/09/2026
Top News

One of AI’s Fiercest Critics Says All the Doom Talk Is ‘Meant to Distract Us’

Palantir is not acceptable to the single English county

Kentucky’s Bitcoin Boom is Over

The AI slop that is flooding their forums has caused cybercriminals to complain.

You might be surprised at how closely the US and China collaborate on AI.

Load More
AI-Trends.Today

Your daily source of AI news and trends. Stay up to date with everything AI and automation!

X (Twitter) Instagram
Top Insights

Sony and Warner Chappell Sue Anthropic Over Claude Lyric Training – Unite.AI

29/08/2026

OpenMythos – A PyTorch Open Source Reconstruction of Claude Mythos, where 770M Parameters match a 1.3B Transformator

19/04/2026
Latest News

Supersonic Labs Releases Julia 1: A 144.3M-Parameter Open Choice Mannequin That Runs on a CPU

27/09/2026

Sarvam AI Releases Saaras V4: A Speech-to-Textual content Mannequin for All 22 Indian Languages and World English

26/09/2026
X (Twitter) Instagram
  • Privacy Policy
  • Contact Us
  • Terms and Conditions
© 2026 AI-Trends.Today

Type above and press Enter to search. Press Esc to cancel.