Open-source AI verification infrastructure for deterministic verification of LLM outputs, tool calls, code, schemas, and agent state before production execution.
-
Updated
Sep 19, 2026 - Python
Open-source AI verification infrastructure for deterministic verification of LLM outputs, tool calls, code, schemas, and agent state before production execution.
Capable, auditable coding that runs fully offline on a 16 GB machine. A verification-first layer (hard test execution, symbolic checking, agentic repair) that takes a local 7B to parity with its 671B teacher on verifiable tasks. MIT, pre-registered, reproducible.
🎓 Free course on deterministic AI verification and AISecOps. Learn fail-closed AI architecture, formal verification, audit integrity, MCP security, and trust-boundary engineering with QWED-AI.
Production-grade epistemic verification for AI agents. Checks semantic compliance, policies, adversarial risks, and reasoning lineage before irreversible actions.
Deterministic reasoning assurance engine for AI agents. Fast (<5ms), zero-cost verification. Best-in-class for arithmetic, logic, and hallucination detection.
VERA: The foundational verification and reliability sandbox for autonomous AI software engineering. Mathematically preventing LLM false-success.
MCP server that gives Claude a review council - other LLMs fact-check responses before you see them
The open-source Fable alternative — a zero-dependency harness that makes ANY LLM verify instead of assume, persist instead of quit, and reuse before reinventing, with a local failover floor you own.
Rage-quit your flaky DB regressions — modern, lightweight, multi-DB regression testing
Claude Code Stop hook that checks an AI assistant's claims against what it actually read this session, using TypeSafe's Jev as the judge
Turn a frozen open-source model into a ~98%-verified, hands-free first-aid assistant — with inference-time compute and a deterministic verifier. Zero training.
Reference implementation and notebook companion repository for “Verified LLM-Assisted Capability- and Skill-Based Process Planning Framework for Modular Plants”.
FinReporting: An Agentic Workflow for Localized Reporting of Cross-Jurisdiction Financial Disclosure, ACL 2026 System Demonstrations.
TriTai 三才 - 零 Token AI 防幻觉引擎,基于太极哲学的 LLM 输出验证系统,集成 WFGY 规则引擎和知识图谱
Verify your Claude Code endpoint really serves GLM-5.2 — tokenizer fingerprint + 1M context probe
Forensic audits for AI coding agents — catches agents that narrate work they never did
Verification for Universal Commerce Protocol (UCP) transactions — Deterministic verification layer for UCP checkouts: catches math, state, and schema errors before payment.
Aether by SF2X — AI trust verification layer. 3-model tribunal that catches LLM hallucinations. 91/100, AUC 1.0. Chrome extension, API, GitHub Action, and public playground.
Deterministic 5-pass LLM verification pipeline — evidence prediction with confidence gating. Provider-agnostic CLI + Python API.
Verify LLM output against your source documents. Catch hallucinations in RAG pipelines and agentic workflows before they reach users.
To associate your repository with the llm-verification topic, visit your repo's landing page and select "manage topics."