RecallRadar Intelligence published the harness and question set for Luna at max, with Luna at max spending so long thinking it ran out of room to answer and four blank responses.
Public source
Publisher name
Public post
more thinking made the models worse. i ran four openai models at every reasoning level on the same 25 questions about fda recall data. 14 configs, 3 reps each, 1,050 cal…
Company
RecallRadar Intelligence
AI Speed. Human-verified.
- Industry
- Research Services
- Location
- Nicoma Park, US
- Company size
- 2–10 employees
About RecallRadar Intelligence
The recall that takes you down didn't start with a recall. It started six months earlier in a warning letter to a supplier you forgot you used. Or a 483 observation on a contract manufacturer two states over. Or an EPR deadline that triggered a packaging change nobody flagged. That data is public. Also buried across FDA, USDA, EPA, CDC, and 50 state portals. None of them talk to each other. Our research agents pull from all of them simultaneously. They cross-reference warning letters, import alerts, 483 observations, enforcement actions, and state rule changes across 5.9 million records. They find the connections a manual search would miss and surface the sources that matter to your operation. Then a human reads every word before it reaches you. No hallucinated citations. No orphaned links. Every source verified, every finding checked against the original regulatory text by an analyst who's worked the floor, the lab, and the land. We tested it. Ranked #20 globally on the DeepResearch Benchmark, ahead of OpenAI, Gemini, and Tavily. AI speed. Human verified. Recalls don't blindside teams watching the right signals. RCRI is taking clients. rcri.io
See moreLatest activity
Latest activity from RecallRadar Intelligence
3 signals
Research & Knowledge
RecallRadar Intelligence published the harness and question set for Sol at high and Sol at xhigh, with Sol at high fabricated 13% of its judgment answers and Sol at xhigh fabricated none and scored about the same.
Research & Knowledge
RecallRadar Intelligence published the harness and question set for Astra at high and Astra at max, with Astra at high achieving 0.938 on judgment and zero fabrications.
Discover more
Similar signals
Similar public activity from other companies.
Research & Knowledge
General Atomics Intelligence
General Atomics Intelligence held a hackathon where 40 participants tested leading AI models through challenges including word games, chess, mazes, visual puzzles, and spatial reasoning.
Research & Knowledge
AlphaSignal
AlphaSignal published a paper showing official scripts went from 67 lines to 374 (5.6x) and the public instruction went from 85 words to 122 (1.4x).
Research & Knowledge
Modern AI Inc.
Modern AI Inc. published its Q3 2026 State of AI Search report measuring 110 brands across ChatGPT, Claude, Gemini, and Perplexity, and ran the same 50 buyer questions through Google search.
Research & Knowledge
Fodda
Fodda ran the Pew Research Center AI Chatbots in Health Report through Fodda, identifying 13 distinct themes and shifts.
Research & Knowledge
Start Solutions AI