
Releases
CollectivIQ Consensus Engine Outperforms Frontier AI Models on Benchmarks
A 96.4% score on the GPQA Diamond benchmark has vaulted Boston-based CollectivIQ ahead of industry-standard frontier models. By querying multiple LLMs simultaneously to identify consensus, the platform achieved results exceeding human PhD baselines and significantly reduced the fabrication errors that typically plague single-model architectures.























