Make AI coding agents architecture-aware: baseline-first, evidence-verified, drift-checked, and safe across long tasks.
-
Updated
Oct 9, 2026 - Python
Make AI coding agents architecture-aware: baseline-first, evidence-verified, drift-checked, and safe across long tasks.
A work-centered runtime for agentic software engineering. Work persists; agents, context, and graphs assemble around it.
随叫随用、证据链驱动的 CUMCM / 数理解、建模学建模 Skills:覆盖题面、复现、图表、审查与提交门禁。
Evidence-driven skill evolution for Hermes Agent — reports, dry-run proposals, candidate search, and guarded apply
面向 AI/CS 科研的多 Skill 实验智能体:助力全自动科研实验,以 Claim 驱动、骨干实验优先、独立结果审计与 Evidence Freeze,把发散的模型能力收束为可追溯、可复现、可收敛的论文证据。
An outcome-first operating framework for reliable AI agent work.
Sanitized Strix 1.2.26 workbench for evidence-driven validation, long-running red-team workflows, and reusable test assets.
A local-first, evidence-driven AI product practice trail—from idea to validation and portfolio.|本地优先、证据驱动的 AI 产品实践路线:从想法走向验证与作品集。
🤖 operational-ai-agents · ⚡ 13 agentes para Claude Code · 🧭 repository-evolution · 🏗️ legacy-modernization · 📡 curriculum-evolution · 🌐 portfolio-publication · 🛡️ security-remediation · 🩺 incident-root-cause · 👤 professional-profile · Contrato + gates humanos + evidencia verificable · 39 evals deterministas · Zero-deps · Cross-platform 🐧🍎🪟
Make any coding model behave like a mythos-class engineer. A behavioural contract, live in-loop correction, and an independent judge that checks the work against your repo — so "done" means proven, not claimed. Self-validating, evidence-driven agent loops for OpenCode.
Chrome-first shopping extension product family with 8 storefront apps, 1 Suite shell, a reviewer-facing release shelf, and a read-only stdio MCP surface.
Evidence-driven Terra/Luna/Sol orchestration for coding agents with durable state, repair, review, and isolated technical-alpha workflows.
Provider-neutral operating standard and reusable assets for AI-assisted collaboration projects
Local-first personal and work operating system with Windows Desktop, full Android mobile, review-first integrations, evidence-backed intelligence and explicit safety controls.
Evidence-first Agent Harness: contracts, validators, State Machine and Evaluation Loop
REI EchoForge: Evidence-driven NPC dialogue for Oblivion/Fallout. Local LLM + Piper TTS + xOBSE adapter. Typed conversation, bounded actions, 100% human review gates. Pre-alpha.
Evidence-driven security and verification control plane for AI coding agents: deny-by-default Policy Engine, sandboxed execution, and a deterministic Decision Engine that computes PASS/FAIL/NEEDS_HUMAN from evidence, never from the agent's own claim.
EvoGuard — Context before merge. Evidence-driven AI code review & provenance tracking. Context-aware code analysis platform that gates AI-assisted Pull Requests against codebase history, contracts, dependencies, tests & security before merge.
An evidence-driven weekly planning Skill that limits preparation and prompts early external validation.
Plataforma determinística e orientada a evidências para diagnóstico e inteligência em engenharia de dados, cobrindo código, arquitetura, Spark, cloud, pipelines, qualidade, segurança, governança, runtime, performance e migrações.
To associate your repository with the evidence-driven topic, visit your repo's landing page and select "manage topics."