Nexus Studio•
Read-only snapshot

Nexus Studio

Sonoma Dark Aurora

Find free and affordable AI models, connect in 3 clicks, and keep your secret keys private.

PULSE Rhythm
#1 Groq LPU
1,250 tps • 4 Rhythms
Active Keys & Apps
0
Public preview: no credential storage
Total AI Models
438
Public catalog snapshot
Catalog Free Tiers
29
Listed free tiers; limits apply
Sovereign Seal
AES-256
Local vault encryption design

PULSE

The Rhythm of AI

Illustrative Market Snapshot. Explore model profiles by speed, value, coding and reasoning; rankings are not live measurements.

142 BPM
Throughput & TTFT
Active Focus:Raw generation speed, sub-second latency, and custom silicon (LPU / TPU / H100).
Showing Top 20 Preview Profiles
🥈#2 Contender
Groq

Groq: Mixtral 8x7B 32k

High-throughput sparse MoE architecture

Rhythm Index
75.2/100
Tokens / Second
480 tps
⚡ Latency:120ms
💰 Cost:Free
Groq LPU Node Cluster
👑 Rhythm Champion
🥇#1 Contender
Groq

Groq: Llama 3.1 8B Instant (~1200 tps)

Blistering sub-100ms real-time generation

Rhythm Index
96.1/100
Tokens / Second
1250 tps
⚡ Latency:98ms
💰 Cost:Free
Groq LPU Node Cluster
🥉#3 Contender
Groq

Groq: Llama 3.3 70B Versatile

Frontier 70B intelligence at instantaneous speed

Rhythm Index
57.5/100
Tokens / Second
280 tps
⚡ Latency:145ms
💰 Cost:Free
Groq LPU Node Cluster
Full Rhythm Standings(#4 through #20)
⚡ Illustrative Performance Profiles
RankModel & ProviderRhythm ScoreThroughputTTFT LatencyPricingMomentumActions
#4
Groq: Google Gemma 2 9B IT
Groq•Groq LPU Node Cluster
57.3pts
260 tps110 msCatalog free tier🎁 Catalog free tier Tier Ready
#5
Groq: Whisper Large v3 (Audio/Speech)
Groq•Groq LPU Node Cluster
57.3pts
260 tps110 msCatalog free tier🎁 Catalog free tier Tier Ready
#6
Groq: Whisper Large v3 Turbo
Groq•Groq LPU Node Cluster
57.3pts
260 tps110 msCatalog free tier🎁 Catalog free tier Tier Ready
#7
Google: Gemini 2.0 Flash (Next-Gen Multimodal)
Google Gemini•Google TPU v5p Pod
53.6pts
240 tps160 msCatalog free tier🔥 +112% viral adoption
#8
Google: Gemini 1.5 Flash (1M Token Window)
Google Gemini•Google TPU v5p Pod
50.1pts
210 tps185 msCatalog free tier+45% WoW traffic
#9
Google: Gemini 1.5 Flash 8B (Sub-second Latency)
Google Gemini•Google TPU v5p Pod
50.1pts
210 tps185 msCatalog free tier+45% WoW traffic
#10
OpenAI: GPT-4o-mini (batch)
OpenRouter•Azure Supercomputer H100
39.6pts
110 tps240 ms$0.07/1M+15% enterprise steady
#11
OpenAI: GPT-4o-mini
OpenRouter•Azure Supercomputer H100
39.6pts
110 tps240 ms$0.15/1M+15% enterprise steady
#12
OpenAI: GPT-4o-mini (2024-07-18)
OpenRouter•Azure Supercomputer H100
39.6pts
110 tps240 ms$0.15/1M+15% enterprise steady
#13
OpenAI: GPT-4o (batch)
OpenRouter•Azure Supercomputer H100
39.6pts
110 tps240 ms$1.25/1M+15% enterprise steady
#14
OpenAI: GPT-4o (2024-11-20)
OpenRouter•Azure Supercomputer H100
39.6pts
110 tps240 ms$2.50/1M+15% enterprise steady
#15
OpenAI: GPT-4o (2024-08-06)
OpenRouter•Azure Supercomputer H100
39.6pts
110 tps240 ms$2.50/1M+15% enterprise steady
#16
OpenAI: GPT-4o
OpenRouter•Azure Supercomputer H100
39.6pts
110 tps240 ms$2.50/1M+15% enterprise steady
#17
OpenAI: GPT-4o (2024-05-13)
OpenRouter•Azure Supercomputer H100
39.6pts
110 tps240 ms$5.00/1M+15% enterprise steady
#18
Qwen2.5 Coder 32B Instruct
OpenRouter•Cloud GPU Cluster
35.7pts
88 tps290 ms$0.66/1M🔥 +76% open source surge
#19
Google: Gemini 1.5 Pro (2M Token Long-Context)
Google Gemini•Google TPU v5p Pod
31.9pts
85 tps380 msCatalog free tier+28% enterprise surge
#20
OpenAI: o3 Mini (batch)
OpenRouter•Azure Supercomputer H100
31.3pts
92 tps410 ms$0.55/1M⚡ New Frontier Entrant
Public preview: no credential storage••