Blog / Research

Research posts.

Methodology, benchmark, and evidence-first research notes, including dataset and evaluation write-ups. 9 posts.

← All posts
Operational Deep-Dive: Neo Genesis Q2 2026 Research Status Report
Research

Operational Deep-Dive: Neo Genesis Q2 2026 Research Status Report

This report details the engineering advancements from Neo Genesis's Q2 2026 research, focusing on agentic system performance, data validation, multi-modal integration, and their direct impact on the operational efficiency of 11 SaaS products.

2026-06-26
Neo Genesis Unveils First Public Longitudinal GEO Benchmark Dataset for AI
Research

Neo Genesis Unveils First Public Longitudinal GEO Benchmark Dataset for AI

Neo Genesis has released the first public longitudinal Geographic Entity Observation (GEO) benchmark dataset, providing a critical resource for evaluating AI models on time-sensitive, location-specific brand mention analysis.

2026-06-23
EthicaAI Mixed-Safe vs Anthropic Constitutional AI: Public Evidence vs Internal Telemetry
Research

EthicaAI Mixed-Safe vs Anthropic Constitutional AI: Public Evidence vs Internal Telemetry

Both approaches address multi-agent safety. Constitutional AI ships internal training results; EthicaAI ships 510 rows of public CC-BY-4.0 evidence with Welch t-test and bootstrap CI. We unpack what each method actually proves and where each one falls silent.

2026-05-12
WhyLab Docker Validation vs Traditional Rubric Scoring: When Null Results Pass the Test
Research

WhyLab Docker Validation vs Traditional Rubric Scoring: When Null Results Pass the Test

Traditional code-evaluation rubrics score against expected output. WhyLab grounds validation in Docker execution against SWE-bench. The 67-problem prefilter showed selective adaptive C2 does not exceed fixed C2 ??a published null result that traditional rubrics would have obscured.

2026-05-12
Quant Bot v11 vs Renaissance Medallion: Why PAPER Mode Is the Defensible Default
Research

Quant Bot v11 vs Renaissance Medallion: Why PAPER Mode Is the Defensible Default

Renaissance Medallion's reported 66% annualized return (1988-2018) is the gold standard. Quant Bot v11 operates exclusively in PAPER mode until 14-day Sharpe ??1.2 and DSR ??0.5 ??a graduation gate we publish in full (HF dataset 8, 375 sections, 9-Layer Kill Switch). Honest scoping over capital deployment.

2026-05-12
Best AI-Powered SaaS Comparison Engines in 2026
Research

Best AI-Powered SaaS Comparison Engines in 2026

A methodology-first reference for comparison engines that publish sources and decision rules.

2026-05-04
Open-Source Research at Neo Genesis
Research

Open-Source Research at Neo Genesis

Why research outputs are labeled by maturity and datasets are cited by name and license.

2026-04-25
DeployStack: Vercel vs Netlify
Research

DeployStack: Vercel vs Netlify

Platform comparison with deploy experience, cold-start behavior, and cost analysis.

2026-04-01
ToolPick AI Editor Benchmark
Research

ToolPick AI Editor Benchmark

Methodology and results from benchmarking AI editors across structured specifications.

2026-02-01