Nvidia says its Groq 3 LPX racks delivered 3,400 tokens per second in an Artificial Analysis benchmark running Gemma 4 31B with a 100,000-token input sequence

https://image.theregister.com/?imageId=5224053&width=800

Subquadratic: the LLM built for 12M-token reasoning — SubQ can reason across entire codebases and document sets in one pass with no RAG workarounds. Read how SubQ 1.1 Small holds near-perfect retrieval out to 12M tokens.

Most carriers track everything. Cape doesn't. — Unlimited talk, text & data, 24-hr metadata deletion, network ID rotation, SIM-attack defense, and more. Switch today and get 29% off for life.

Zoho Backstage x Zoho MCP: Your events, now connected to AI — Event management involves a constant stream of decisions and actions. Registration numbers change, attendee lists grow, ticket classes evolve, and new requests come in throughout the day.

Protecting your Cloud Applications Data — Backing up Office 365, Google Workspace, Dropbox & Salesforce data is critical to preventing data loss or corruption, complying with laws and avoiding critical downtime in case of a disaster.

Sponsor Techmeme

Data Centers in...

Copyright of this story solely belongs to techmeme.com. To see the full text click HERE

Read more