Open-source AI verification infrastructure for deterministic verification of LLM outputs, tool calls, code, schemas, and agent state before production execution.
-
Updated
Oct 1, 2026 - Python
Open-source AI verification infrastructure for deterministic verification of LLM outputs, tool calls, code, schemas, and agent state before production execution.
Capable, auditable coding that runs fully offline on a 16 GB machine. A verification-first layer (hard test execution, symbolic checking, agentic repair) that takes a local 7B to parity with its 671B teacher on verifiable tasks. MIT, pre-registered, reproducible.
🎓 Free course on deterministic AI verification and AISecOps. Learn fail-closed AI architecture, formal verification, audit integrity, MCP security, and trust-boundary engineering with QWED-AI.
Production-grade epistemic verification for AI agents. Checks semantic compliance, policies, adversarial risks, and reasoning lineage before irreversible actions.
Deterministic reasoning assurance engine for AI agents. Fast (<5ms), zero-cost verification. Best-in-class for arithmetic, logic, and hallucination detection.
VERA: The foundational verification and reliability sandbox for autonomous AI software engineering. Mathematically preventing LLM false-success.
Rage-quit your flaky DB regressions — modern, lightweight, multi-DB regression testing
MCP server that gives Claude a review council - other LLMs fact-check responses before you see them
The open-source Fable alternative — a zero-dependency harness that makes ANY LLM verify instead of assume, persist instead of quit, and reuse before reinventing, with a local failover floor you own.
Claude Code Stop hook that checks an AI assistant's claims against what it actually read this session, using TypeSafe's Jev as the judge
Turn a frozen open-source model into a ~98%-verified, hands-free first-aid assistant — with inference-time compute and a deterministic verifier. Zero training.
TriTai 三才 - 零 Token AI 防幻觉引擎,基于太极哲学的 LLM 输出验证系统,集成 WFGY 规则引擎和知识图谱
Reference implementation and notebook companion repository for “Verified LLM-Assisted Capability- and Skill-Based Process Planning Framework for Modular Plants”.
FinReporting: An Agentic Workflow for Localized Reporting of Cross-Jurisdiction Financial Disclosure, ACL 2026 System Demonstrations.
Verify your Claude Code endpoint really serves GLM-5.2 — tokenizer fingerprint + 1M context probe
Forensic audits for AI coding agents — catches agents that narrate work they never did
Multi-provider coding-agent orchestration in a verified loop: a critic on a different model signs off before work ships. Claude Code, Codex, Gemini, aider, Copilot, opencode, local.
Aether by SF2X — AI trust verification layer. 3-model tribunal that catches LLM hallucinations. 91/100, AUC 1.0. Chrome extension, API, GitHub Action, and public playground.
Verification for Universal Commerce Protocol (UCP) transactions — Deterministic verification layer for UCP checkouts: catches math, state, and schema errors before payment.
Verify LLM output against your source documents. Catch hallucinations in RAG pipelines and agentic workflows before they reach users.
To associate your repository with the llm-verification topic, visit your repo's landing page and select "manage topics."