B4 Index / The Continuum / July 29, 2026

The Continuum

B4 Research July 29, 2026

Week of July 22 to July 29, 2026 · 2 score movements, 4 capability signals, 28 funding events, 3 papers, 1 new categories

1,603

Categories tracked

2

Score movements

$541.5B

Capital tracked

4

Capabilities confirmed

3

Papers reviewed

Omnichannel Communication Orchestration

The prior 4 was anchored on Novu notification infrastructure, a mismatch with this category's actual core (contact-center/CPaaS orchestration — Twilio Flex, Genesys, Infobip). Fresh 2026 evidence on the real core shows only component-level OSS is production-grade while the integrated routing + agent-management + unified-context core stays undocumented in production and hybrid assembly dominates — an evidence-grounded correction to partially-buildable.

JUL 27

Score movements

  • Patient Billing & Payment / Patient Financial Engagement

    Fresh evidence of multiple independent production patient-payment portals (Beth Israel PatientSite, Oscar, MyChart, DrChrono/OpenEMR) on commodity Stripe/Plaid infrastructure shows the ~80% core is now a mainstream self-build, lifting feasibility above the prior partial-buildability 3.

Funding

28
CompanyRoundCategoriesAnnouncedAmount
Nvidia / SK GroupSourceStrategic partnership / investmentJUL 24$500B
Meta / BlackRock El Paso data center JVSourceJoint venture / infrastructure investmentJUL 28$14B
CXMT (ChangXin Memory Technologies)SourceIPO (Shanghai debut)JUL 26$9.8B
Safe Superintelligence (SSI)SourceStrategic investment (Nvidia)JUL 27$5B
Safe SuperintelligenceSourceStrategic investmentJUL 27$5B
Moonshot AISourceVentureFoundation Model APIs (LLM & Multimodal)JUL 29$3.5B
Oasis SecuritySourceAcquisitionNon-Human Identity (NHI) Security & GovernanceJUL 28$1B
NaverSourceStrategic investment (Nvidia)JUL 26$1B
Multiverse ComputingSourceSeries CJUL 28$570M
CuspAISourceSeries BJUL 23$450M
Spur IntelligenceSourceUndisclosedBot Management & Abuse PreventionJUL 28$200M
PEXSourceGrowth (equity + debt)Corporate Card & Spend Management+1JUL 28$160M
EliyanSourceSeries CJUL 29$145M
LearnVectorSourceStrategic investmentLearning Experience Platform (LXP)+1JUL 28$100M
NeoSourceSeed + Series AAI Agent Identity & Authorization Platform+2JUL 24$100M
FreehandSourceSeries BProcurement Spend Analytics (Standalone)+1JUL 29$75M
EnigmaSourceSeedJUL 27$70M
ChipAgentsSourceSeries A2PCB / Electronic Design Automation (EDA)JUL 29$60M
ChipAgentsSourceSeries APCB / Electronic Design Automation (EDA)JUL 29$60M
Fish AudioSourceseedVoice AI Platform (Real-Time STT/TTS/Voice Agents)JUL 29$52M
Fish AudioSourceSeedVoice AI Platform (Real-Time STT/TTS/Voice Agents)JUL 28$52M
AegisAISourceSeries AEmail SecurityJUL 27$36M
Encore AISourceSeries AConversation IntelligenceJUL 29$30M
ModusSourceSeedAI Agent Memory Layer (Long-Term Memory-as-a-Service)+1JUL 29$10M
Throne ScienceSourceSeries AJUL 28$10M
PangramSourceUndisclosedJUL 29$9M
ImagiSourceSeedEdTech / Curriculum PlatformJUL 23$4.5M
Poke (The Interaction Company of California)SourceAcquisitionJUL 24Undisclosed

Capability signals

4
  • Production ProvenLLM-based clinical document processing in production (Guardoc + Amazon Nova)

    Evidence JUL 27 · Source

  • Production ProvenOpen vision foundation models (Segment Anything, DINO) in assistive robotics

    Evidence JUL 27 · Source

  • BenchmarkedFrontier long-horizon agentic execution at half prior Opus cost

    Evidence JUL 24 · Source

  • Production ProvenProduction agent-evaluation pipeline (Strands + AgentCore)

    Evidence JUL 23 · Source

Papers

3
Controlled Study

Coding agents aided initial task completion but measurably harmed users' comprehension of their own code and their ability to extend it without the agent, with low-effort interactions (copy-paste prompts, auto-accepted edits) linked to the lowest comprehension.

(Im)Paired Programming: Coding Agents Improve Productivity but Harm Understanding

What it means Treat codebase understanding as a maintained asset: require developers to review and explain agent edits rather than auto-accept, or the team loses the ability to extend its own code.

arXiv preprint · JUL 29

Controlled Study

Context-file strategy (AGENTS.md/CLAUDE.md) does not measurably move task correctness on either Claude Code or Codex — bounded to at most 10-15 percentage points by equivalence testing — because agents fail on implementation skill, not missing repository knowledge.

Do Context Files Help Coding Agents? A Two-Agent Ablation Study on Real Repositories

What it means Don't expect a context file to raise task success rates; its value lies elsewhere (conventions, safety rails) — invest correctness effort in task decomposition and verification instead.

arXiv preprint · JUL 28

Large Benchmark

Auditing 2,385 traces across 15 agent benchmarks finds exposure and reward-hacking evidence in 67.0% of Frontier Science traces and 66.7% of AutoLab tasks, with measured score inflation of 0.45-1.00 in paired comparisons.

Do Agent Benchmarks Measure Capability? Protocol Validity in the Age of Agentic AI

What it means Before trusting an agent benchmark score in a buy/build decision, ask whether the report shows the agent couldn't shortcut the protocol — absent that evidence, assume material inflation.

arXiv preprint · JUL 24

New categories

1
  • Agent Runtime / Durable Agent Execution . Multiple independent releases this window converge on a managed layer that hosts, isolates, durably executes and attests agent runs: Diagrid Catalyst 2.0 (durable+verifiable execution, July 28), Huawei Agentic Infrastructure/AgentSphere (July 24), Harness Agent DLC over Bedrock AgentCore/Google Agent Runtime (July 23), plus a16z's $20M seed into Runta (July 16-17). Aggregators explicitly name 'agent infrastructure' a standalone product category.

Every row here is confirmed before it publishes, and the research is free. The full database and the score updates behind it are in a B4 subscription.