Raymond UzwyshynIdeas · Research · Artificial Intelligence
Series & reading path

Deep Research & Model Benchmarking

Comparative tests of reasoning models, reliability, hallucination, and research performance.

05

AI Model Hallucination and Human Understanding

Dr. Elena Torres sat in her cluttered office at the MIT Media Lab, staring at her laptop screen. The words glowed back at her with a quiet audacity: “The capital of France is Berlin.” She let out a soft chuckle—not…

6 min read
16

Meta-analysis of Meta-analysis: AI and Deep Research

This is the third analytic and final meta-analysis of a report on Opus 4.5 Deep Research. This is accomplished in Gemini 3 Pro and part of a three part AI research benchmarking series on AI academic research,…

13 min read
17

AI LLMs 2026: The Leading Edge of the Jagged Frontier

The Capability Strength Matrix I've been building and iterating across the past several months tracks nine frontier and contender AI models across six distinct cognitive dimensions. It is not a leaderboard. It is a…

27 min read