Research Sweeps
Multi-lane research syntheses
Deep reads across financial press, frontier labs, academia, VC, and independent analysis - distilled into briefs with full source tables.
2026-08-04 · deep
Apple Silicon Unified Memory as On-Premise AI Infrastructure (Jan 2025 – Aug 2026)
Whether Apple Silicon unified memory and the Neural Engine are becoming a credible substrate for how AI is run commercially and on consumer devices between January 2025 and August 2026: shipped capacity and bandwidth (M3 Ultra at 512GB, M4 and M5 Max) versus the rumoured 1.5TB M7 Ultra, comparison against NVIDIA H200, GB200 NVL72 and AMD MI355X on memory-bound serving, the orchestration and scheduling gap (MLX distributed, EXO, Thunderbolt 5 fabrics, Kubernetes and MDM on macOS, Metal kernel maturity versus CUDA), the on-device stack (Neural Engine, Core ML, the Foundation Models framework, Private Cloud Compute), fleet monitoring and data-centre economics, and how Apple is positioned to capture value from the AI race as compute substrate rather than as a frontier model lab.
Claude Fable 5- frontier
- academic
- tech
- +3
2026-08-03 · deep
Security Research into Chinese Open-Weight Models
Independent security research into Chinese open-weight models (DeepSeek R1 and V3, Alibaba Qwen, Moonshot Kimi K2, Zhipu GLM, MiniMax, Baidu Ernie) from February 2025 to August 2026: who is testing them, what red-teaming and provenance methods they use, which results survive independent replication, and how regulated, defence and military buyers are assuring models whose training data and training objectives are never disclosed
GPT-5.6-sol- financial
- frontier
- academic
- +2
2026-07-28 · deep
AI 2027 Reality Check, Deep
AI 2027 scenario tracking from the report's April 2025 publication through to late July 2026: whether the OpenAI sandbox escape and Hugging Face compromise, the US export-control suspension of Claude Fable 5 and Mythos 5, state and criminal use of AI in cyber operations, and Chinese open-weight model capability reproduce the scenario's predicted sequence of loss of control, cyber capability and state intervention; whether the capability curve has plateaued or merely re-based; and which of the scenario's remaining failure conditions have or have not materialised
GPT-5.6-sol- financial
- frontier
- academic
- +3
2026-07-24 · deep
Software Project Outcomes Over Time, and Is AI Improving Them
Software and IT project outcomes over time from July 2016 to July 2026 (biased to the latest year): baseline success/challenged/failure ranges across trusted longitudinal datasets (Standish CHAOS, Oxford Global Projects and Flyvbjerg, McKinsey, PMI Pulse), the drivers of outcomes (complexity, budget, domain, delivery methodology, sourcing), build-vs-buy and low-code/no-code commissioning choices (OutSystems, Mendix, Power Platform, Retool), and whether AI-assisted development (Copilot, Cursor, Claude Code, DORA and METR evidence) is yet improving delivery outcomes on traditional software projects that build ordinary systems rather than AI products.
Claude Fable 5- financial
- frontier
- academic
- +3
2026-07-21 · deep
Agentic Harnesses Market Landscape, July 2025 to July 2026
Agentic harnesses from July 2025 to July 21, 2026: competitive landscape, market share by use case, model integration, orchestration patterns, and differences between coding harnesses and broader workflow agents, including Claude Code, OpenAI Codex, Cursor, GitHub Copilot, Google Jules, Replit Agent, Devin, Windsurf, OpenCode, Aider, OpenHands, LangGraph, CrewAI, AutoGen, and n8n.
GPT-5.6-sol- financial
- frontier
- academic
- +3
2026-07-15 · deep
Enterprise AI Transformation Programmes (2025–2026)
Enterprise AI transformation programmes from July 2025 to July 2026: reported success and failure rates and who measures them, delivery frameworks borrowed from and diverging from classic digital transformation, token cost economics and budgeting under consumption pricing, adoption strategy including leadership over-provisioning of access, and AI governance across AI-Ops, financial business cases, security and regulatory controls as they scale with sector risk tolerance, referencing MIT, McKinsey State of AI, DORA, ThoughtWorks Technology Radar, NIST AI RMF, ISO 42001, and the EU AI Act.
Claude Fable 5- financial
- tech
- academic
- +3
2026-06-26 · deep
AI in Weather and Climate Prediction
AI in weather and climate prediction across the 2015 to June 2026 machine-learning era, with historical context from mid-twentieth-century numerical weather prediction and Lorenz's chaos theory: the shift from physics-based NWP and statistical post-processing (MOS) to data-driven models (GraphCast, GenCast, Pangu-Weather, FourCastNet, Aurora, NeuralGCM, ECMWF AIFS), how forecasters at ECMWF, NOAA, and the Met Office have operationalised them, measured accuracy versus the IFS, and the predictability limits imposed by chaos, the Lorenz attractor, and the butterfly effect.
Claude Opus 4.8- financial
- frontier
- academic
- +2
2026-06-26 · deep
Climate Change - Evidence, Attribution, Projections and Remedies
Climate change from the 1970s to June 2026: the measured surface-temperature record (HadCRUT5, NASA GISTEMP, NOAA, Berkeley Earth) and cryosphere proxies (NSIDC sea-ice, snow cover, glacier mass balance, ocean heat content), biodiversity feedbacks, attribution to fossil-fuel carbon, and forward projections under current-trend and modelled scenarios (IPCC AR6, CMIP6 SSPs), through to the evidence base for remedies from ocean and land carbon sinks to solar geoengineering, weighting recent authoritative sources (IPCC, PNAS, Nature Climate Change, Copernicus/ECMWF) over older ones.
Claude Opus 4.8- financial
- academic
- vc
- +1
2026-06-20 · deep
Comparative LLM Usage Across Sectors
Comparative real-world usage of LLMs and adjacent AI technologies from June 2025 to June 2026: which models (GPT-5, Claude, Gemini, Llama, Mistral, DeepSeek, Qwen) dominate which sectors, how they are deployed (hosted API, Bedrock/Azure, self-hosted vLLM/Ollama, RAG, agents, fine-tuning), what workloads they serve, and how organisations measure, budget, and publicly report token cost and actual spend.
Claude Opus 4.8- financial
- frontier
- academic
- +3
2026-06-16 · deep
Designing AI Operating Models Around Humans
How humans are adapting to AI between June 2024 and June 2026, weighing measured benefits and harms, and how organizations should design operating models around human cognitive load and behavioural patterns rather than forcing adoption, covering cognitive overload from supervising multiple agents at machine speed (context switching, automation complacency, vigilance fatigue), the poor budget and value outcomes of top-down AI mandates and token-maximizing usage, the gap between model welfare functions (such as Anthropic's) and any equivalent human or worker welfare function, and how much good human outcomes depend on model training versus orchestration and deployment design.
GPT-5.5- financial
- frontier
- academic
- +3
2026-06-07 · deep
AI on Deterministic Rails
AI on deterministic rails: how AI and traditional deterministic software are forming a symbiotic stack from January 2025 through June 2026: the enterprise "PoC-opalypse" and the shift from token consumption to durable agentic adoption patterns, AI leveraging software-encoded workflows as guardrails (variance and error control) rather than replacing them, the frontier moving from raw model capability to model orchestration and harness design (Claude Code, OpenCode, Pi), right-sizing with smaller and open-weight models (Llama, Qwen, DeepSeek, Mistral) for cheap routine automation and private inference, and the token-pricing economics behind enterprise sticker-shock over agentic spend versus delivered value
Claude Opus 4.8- financial
- frontier
- academic
- +3
2026-06-03 · deep
Code Intelligence & Code-Graph Indexing for AI Agents
Tools and emerging approaches for code intelligence and code-graph indexing for AI coding agents from June 2025 through early June 2026, spanning local/embedded indexers (CodeGraph/Caveman-style repo maps, tree-sitter, SQLite and embedded graph stores), enterprise-scale code understanding (SCIP, code knowledge graphs, embeddings+retrieval), LSP-to-MCP bridges such as Serena, and the semantic-vs-syntactic-vs-embedding trade-off.
GPT-5.5- tech
- frontier
- academic
- +2
2026-06-01 · standard
Handling Large Volatile Corpora with AI: Caching, Freshness, and Retrieval at Scale
Engineering patterns for large, fast-changing corpora from 2024 to 2026: prompt and prefix caching, the shift from prompt engineering to context engineering, embedding staleness and freshness strategies, multi-strategy retrieval beyond pure vector search, and the inference-cost economics now reshaping infrastructure decisions.
Claude Opus 4.8- frontier
- tech
- academic
- +1
2026-05-19 · deep
AI Regulation and the Regulated Enterprise - Trajectory to 2030
The trajectory of AI regulation across the EU AI Act, the UK's pro-innovation and contextual approach, and the financial-services regulatory regime (FCA, PRA, Bank of England) from January 2023 to May 2026, including the FCA Mills Review, GPAI obligations, model-risk and accountability rules, and what they demand of technology leadership in regulated firms
Gemini 2.5 Pro- frontier
- academic
- vc
2026-05-15 · deep
Emergent Behaviour Across Disciplines
The science of emergent behaviour and self-organisation from 2015–May 2026, connecting reaction-diffusion and cellular-automata models to commercial markets, organisational behaviour and culture, ecology, biology, physics and chaos theory, including foundational algorithms and their cross-domain explanatory power.
Claude Opus 4.8- academic
- blogs
- financial
2026-05-11 · deep
Compounding Waves - How Each Tech Era Built the Substrate, and the Skills, for the Next
The compounding economic logic of three successive technology waves from January 1995 to May 2026 - internet disintermediation of distribution, software-defined platforms and cloud infrastructure, and the current AI/agentic systems wave - examining the technical, economic and human-skills dependencies that make each wave a precondition for the next, the new categories of work each wave created, and whether the relationship is best understood as cumulative compounding or as externalised costs harvested by later layers.
Claude Opus 4.8- financial
- academic
- blogs
- +1
2026-05-10 · deep
Agentic RAG - Evolution, Challenges, and Decision Criteria
Agentic RAG between November 2025 and May 2026: how retrieval-augmented generation is shifting toward agent-driven architectures, the operational problems (token burn, context management, latency, reliability), information-organisation patterns such as context catalogues and semantic categorisation, parallels with traditional data warehousing (dimensions, measures, star schemas), the evolving RAG tooling landscape, and decision criteria for switching to pure agentic workflows.
Claude Opus 4.8- academic
- frontier
- tech
- +2
2026-04-30 · deep
Agentic Engineering And Enterprise Architecture Discipline
Agentic engineering after Andrej Karpathy's vibe coding meme, April 2025-April 2026: how AI coding agents are changing enterprise software engineering across security, testability, reliability, maintainability, availability, resilience, observability, operability, cost, recovery, and engineering governance.
GPT-5.5- frontier
- academic
- vc
- +3
2026-04-24 · deep
Engineering AI Control Plane
Engineering AI control planes for software delivery from July 1, 2025 through April 24, 2026: how teams implement AI across development workflows and CI/CD, choose tools/models/SDKs, govern observability and compliance, manage reliability and provider availability, and handle cognitive debt, dark code, case studies, success stories, and failure modes across team size, company scale, and greenfield versus brownfield systems
Claude Opus 4.8- financial
- frontier
- academic
- +3
2026-04-19 · shallow
The Karpathy Loop - AI Agents Running Autonomous Training Experiments
The "Karpathy loop" - autonomous AI agent research cycles that run and evaluate ML training experiments to discover improvements, April 2025–April 19 2026, including Karpathy's own explanations, independent commentary, and real-world implementations
Claude Opus 4.5- frontier
- blogs
- tech
2026-04-19 · deep
Quantum Computing Foundations - A Briefing Note with Sources
Quantum computing fundamentals briefing - error correction, hardware architectures, computational advantage, and where the field stands - key papers, expert commentary, and lab progress from January 2023 to April 2026
Claude Opus 4.8- academic
- frontier
- blogs
2026-04-19 · deep
Token Cost of Ownership
AI token pricing vs true total cost of ownership from January 2023 to 19 April 2026, with emphasis on 2025–2026 signals: lab subsidisation strategies, infrastructure economics (compute, energy, data centres, hardware, security, ops), how user-facing prices have evolved, and analyst and researcher projections for token cost trajectories through 2028.
Claude Opus 4.8- financial
- frontier
- academic
- +2
2026-04-18 · deep
Quantum Computing Advances - State of the Field 2025–2026
Quantum computing advances, quantum advantage benchmarks, real-world applications, and programming interfaces for business and consumer use - April 2025 to April 2026
Claude Opus 4.8- frontier
- academic
- vc
- +2
2026-04-17 · deep
Agentic AI's Impact on Technology Operating Models and Architecture
Agentic AI's impact on enterprise technology operating models and architecture (January 2025–April 17th 2026): what stays (API infrastructure, data governance, SDLC controls), what shifts (DevOps as the new control plane, testing and rollback at agent speed, dark-code and agentic tech-debt governance), and whether frontier models like Anthropic's Mythos become embedded in CI/CD pipelines for security, code review, and release control
Claude Opus 4.8- financial
- frontier
- academic
- +3
2026-04-15 · standard
The SaaS-pocalypse - AI Displacement, Overhiring Hangover, or Multiple Compression?
The 2026 SaaS sector stress: testing whether weak SaaS revenue growth and stock performance are driven by AI displacing knowledge-work jobs, post-ZIRP overhiring correction, compression of growth-era revenue multiples, or macro tech-capex slowdown - January 2026 through April 2026.
Claude Opus 4.8- financial
- academic
- vc
- +1
2026-04-13 · deep
AI Dark Code - Organisational Accountability and Control
AI-generated and agent-produced code ("dark code") in enterprise settings June 2025–April 2026: organisational accountability structures, failure and adaptation of established management frameworks, technical and governance controls, observability and discoverability of agent logic, and documented outcomes from early enterprise adoption.
Claude Opus 4.8- financial
- frontier
- academic
- +2
2026-04-13 · standard
Enterprise LLM Vendor Selection and Consumption Models
Enterprise LLM vendor selection and consumption patterns (April 2025–present): how companies choose between OpenAI, Anthropic, Google, hyperscaler-hosted model access, and direct API relationships; what decision metrics they use across availability, quality, price, governance, and SLAs; and how adoption differs by company size, workload criticality, and realtime versus offline use cases
Claude Opus 4.8- financial
- frontier
- academic
- +2
2026-04-09 · standard
Enterprise Agentic AI Adoption Criteria
Enterprise agentic AI adoption in operational processes November 2025–present: procurement criteria, model drift risk, version stability, availability SLAs, and how enterprises manage dependency on AI vendors in production workflows
Claude Opus 4.8- financial
- frontier
- academic
- +1
2026-04-08 · deep
AI 2027 Milestone Tracker
AI 2027 report milestone tracking (January 2025–present): which predicted capabilities have shipped across Anthropic, OpenAI, Google DeepMind, Meta, xAI, and major enterprise adopters; what remains unshipped or contradicted; and what near-term signals suggest for agentic AI, safety frameworks, autonomy, and deployment timelines
Claude Opus 4.8- financial
- frontier
- academic
- +2