Oxmiq Labs reported that HydraFusion delivered a 67% cost reduction and a 4.9 percentage point improvement in verified task quality on TerminalBench 2.1 versus Claude Opus 5.
Public source
Publisher name
Public post
๐๐ถ๐๐๐๐ฏ ๐๐ผ๐ฝ๐ถ๐น๐ผ๐ ๐ท๐๐๐ ๐๐ต๐ถ๐ฝ๐ฝ๐ฒ๐ฑ ๐๐๐ฑ๐ฟ๐ฎ๐๐๐๐ถ๐ผ๐ป ๐ฎ๐ ๐ฎ ๐ฟ๐ฒ๐๐ฒ๐ฎ๐ฟ๐ฐ๐ต ๐ฝ๐ฟ๐ฒ๐๐ถ๐ฒ๐. ๐๐ ๐ฟ๐ผ๐๐๐ฒ๐ ๐ฐ๐ผ๐ฑ๐ถ๐ป๐ด ๐๐ฎ๐๐ธ๐ ๐ฎ๐ฐ๐ฟ๐ผ๐๏ฟฝโฆ
Company
Oxmiq Labs
Re-Architecting the GPU Stack: From Atoms to Agentsโข
- Industry
- Technology, Information and Internet
- Location
- Campbell, US
- Company size
- 51โ200 employees
About Oxmiq Labs
Founded by Raja Koduri, OXMIQโข is rearchitecting the GPU full stack from Atoms to Agents to meet the demands of next-generation gaming, graphics and multimodal AI. The company develops licensable GPU software and hardware IP that balances flexibility and performance by integrating breakthrough technologiesโincluding nano agents in silicon based on RISC-V, near- and in-memory computing, advanced light transport, and other innovations. OXMIQโs architecture is designed to scale seamlessly from physical AI devices all the way to Data Center scale. See oxmiq.ai for more info.
See moreLatest activity
Latest activity from Oxmiq Labs
19 signals
Products & Services
Oxmiq Labs announced that Nebius and Palantir are offering a bundled stack including infrastructure and operational intelligence tools designed to meet data residency and security requirements.
Partnerships
Oxmiq Labs announced that Palantir brings the software layer.
Partnerships
Oxmiq Labs announced that Nebius brings NVIDIA-based GPU clusters across the U.S. and Europe.
Discover more
Similar signals
Similar public activity from other companies.
Research & Knowledge
Microsoft
Microsoft announced that in offline testing on TerminalBench 2.1, HydraFusion improved verified task quality by 4.9 percentage points over the evaluated Claude Opus 5 baseline, at an estimated 67% lower cost.
Research & Knowledge
Yantrion Inc
Yantrion Inc measured 975.5 aggregate decode tokens per second across 56 concurrent requests on one 8ร AMD Instinct MI350X node using Kimi-K3.
Research & Knowledge
Workfabric AI
Workfabric AI tested that in an actual live production task, where the precise context shipped vs generic context results in 50% more accurate results in the former while also lowering token costs.
Research & Knowledge
Concentrate AI
Concentrate AI reports that 1 billion tokens on GLM-5.3-Flash costs ~$35, which is 26x cheaper than GPT-5 High.
Research & Knowledge
NVIDIA