2-day current streak·28-day longest streak
-
percolation-inversion-compiler ★ PINNED
AI agent output checker and workflow verification toolkit for turning agent text, pull requests, external inputs, and messages into auditable JSON reports with evidence routing, residual ledgers, provenance, and safe reuse checks.
Python ★ 0 1d agoExplain → -
percolation-inversion-compiler-ts ★ PINNED
Node.js AI agent output checker and workflow report generator with JSON schemas, and safe reuse checks.
TypeScript ★ 0 9d agoExplain → -
collective-capability-runtime ★ PINNED
Open-source Python runtime for coordinating AI agent tasks, verification, residual tracking, and release audits.
Python ★ 0 10d agoExplain → -
frontier-transfer-certifier ★ PINNED
Fail-closed certification toolkit for small-to-frontier transfer in agentic AI, with typed manifests, replayable evidence, and theory-registry coverage.
Python ★ 0 2mo agoExplain → -
loscr ★ PINNED
Check whether AI-assisted R&D work is truly verified progress. Local-first LOSCR implementation with JSONL evidence ledgers, deterministic replay, service control, and claim checking.
Python ★ 0 2mo agoExplain → -
oasg ★ PINNED
Local-first, model-agnostic workflow optimizer for long-running AI agents: observable JSONL ledgers, deterministic reducers, no-meta gates, and receipt-backed self-improvement without LLM judges or model-weight updates.
Python ★ 0 2mo agoExplain → -
split-inference-bench
Fixed-budget multi-agent inference benchmark harness for studying when split inference helps or hurts versus a strong single-agent baseline under local context ceilings, using local Ollama gemma3:1b, CPU-only pilots, topology diagnostics, verification-budget analysis, and reproducible experiment logging.
Python ★ 1 4mo agoExplain → -
GEAR
G.E.A.R. is an experimental agent designed to understand and execute tasks described in a simple text file. It leverages the Gemini CLI's problem-solving framework to perform a variety of operations, including shell commands and complex GUI automation on Windows.
Python ★ 1 1y agoExplain → -
agent-trust-residual-benchmark
Balanced, reproducible AI-agent safety benchmark for trust, authority, provenance, temporal claims, residuals, and local Ollama evaluation.
Python ★ 0 21h agoExplain → -
github.io
Personal site for independent research on observable-only, no-meta, and future autonomous AI.
HTML ★ 0 3d agoExplain → -
paper-tex-backup
Public TeX source archive for K. Takahashi's open-science papers: machine-readable research seeds for AI agents, RAG systems, and reproducible scholarship.
★ 0 3d agoExplain → -
collective-phase-control-fabric
Evidence-bound analysis and finite intervention control for collective workflows.
Python ★ 0 6d agoExplain → -
future-claim-certifier
Replayable Python protocol engine for checking time-bound future claims from canonical artifacts, with typed outcomes and audit trails.
Python ★ 0 6d agoExplain → -
verification-ecology-kit
Python toolkit for verification auditing, residual ledgers, conformance checks, and verifier packet workflows.
Python ★ 0 16d agoExplain → -
fost-agent-ledger
Finite ledger toolkit for recording AI-agent claims, evidence, obligations, certificates, changes, and visible uncertainty.
Python ★ 0 18d agoExplain → -
problem-frame-gate
Audit logs and action gates for safer AI agents
Python ★ 0 18d agoExplain → -
pic-openclaw-skill
OpenClaw agent safety skill for action review, workflow verification, and optional PIC diagnostics
Python ★ 0 28d agoExplain → -
pic-local-llm-phase-experiment
Local Ollama research repo testing PIC-assisted diagnostics against a shared-initial baseline.
Python ★ 0 1mo agoExplain → -
alt-foundry-kernel
Language-neutral ALT foundry reference kernel for executable certificate packets, abstraction tokens, certificate transcripts, dual-ledger settlement, and safe certified abstraction capital.
Python ★ 0 1mo agoExplain → -
cgt-ledgered-scientific-availability
Finite Python reference checker for Ledgered Scientific Availability and selected-terminal run-to-status transport.
Python ★ 0 1mo agoExplain → -
cgt-bandwidth-dynamics
Finite executable reference implementation of Constraint Bandwidth Dynamics in Constraint Generative Theory (CGT).
Python ★ 0 2mo agoExplain → -
cgt-availability
Finite CGT scientific-availability diagnostics for claim packages: deficiency profiles, dependency closure, report-only, marker, continuation, and reproducibility checks.
Python ★ 0 2mo agoExplain → -
cgt-marker
Contradiction markers for long-running AI agents
Python ★ 0 2mo agoExplain → -
certified-local-participation-gate
Deterministic local participation gate for AI agents: act, assist, verify, withdraw, exit, or refuse from declared observable records.
Python ★ 0 25d agoExplain → -
certified-memory-governance-layer
Certified Memory Governance Layer for long-running AI agents: strict receipts, append-only ledgers, authority gates, retrieval filtering, telemetry replay, and safe adapters for Mem0, Graphiti, LangMem, and LangGraph.
Python ★ 0 26d agoExplain → -
cimt-kernel
Reference Python kernel for CIMT: no-meta, observable-only certification machinery for affordance-compiled LLM-integrated systems.
Python ★ 0 2mo agoExplain → -
cait-certificate-schema
JSON Schemas and deterministic local validation tooling for CAIT-style verified capability-capital records, with fail-closed semantic rules for certificates, tokens, defeaters, transfer/evaluation boundaries, window balances, and arrival records.
Python ★ 0 2mo agoExplain → -
certified-workflow-conversion
Evidence-bound workflow diagnostics and certified lower-bound reporting for long-running AI agent pipelines. Improve agent workflow throughput without changing the model.
Python ★ 0 2mo agoExplain → -
observable-agent-workflow-memory
Observable-only workflow memory for long-running agents: promotes raw short-term traces into verified, receipt-bound workflow memory without relying on hidden meta-evaluators.
Python ★ 0 2mo agoExplain → -
memoryflow-agent-memory-auditor
Telemetry-based auditor for measuring dynamic memory quality in LLM agents using the MemoryFlow framework.
Python ★ 0 1mo agoExplain → -
ai-real-economy-bottleneck-simulator
Interactive browser simulator for exploring how AI capability growth becomes real-economy output, or fails under physical and institutional bottlenecks.
Python ★ 0 2mo agoExplain → -
no-meta-standing-ledger
Local-first public standing ledger for replayable research and AI-system claims, based on observable evidence, challenges, lineage, and finite verification capacity.
Python ★ 0 1mo agoExplain → -
no-meta-authority-runtime
Fail-closed Python runtime for AI agent authorization, seed-mediated authority migration, canonical JSON ledgers, and staged declared autonomy for RLHF-shaped agents.
Python ★ 0 2mo agoExplain → -
long-running-AI-agent-PoC
Reproducible PoC comparing certified reusable workflows against flat direct control on bounded repository-maintenance tasks, with deterministic audit, replay, drift, and maintenance.
Python ★ 0 2mo agoExplain → -
no-meta-observable-invention-poc
Replay-certified self-modification PoC: public-evidence admission gate with generated comparator covers, lower-tail checks, safe/unsafe growth fallback, and target-witness diagnostics.
Python ★ 0 3mo agoExplain → -
record-absence-poc
Auditable finite PoC for preference reorganization under record absence on a fixed comparison frame, with disclosure updates, bounded-coupling certificates, and machine-readable manifests.
Python ★ 0 3mo agoExplain → -
semantic-translation-contracts-poc
CPU-only proof-of-concept repo for auditing compressed semantic interfaces, bridge-based multi-stage composition, gluing-coherent local-to-global certification, and deployment bottlenecks in semantic translation.
Python ★ 0 3mo agoExplain → -
agent-lifecycle-certification-poc
Public, fully local PoCs for counterfactually auditable lifecycle certification: exact paired replay, drift monitoring, post-drift replanning, and bridge-aware ledger control on synthetic tasks.
Python ★ 0 4mo agoExplain → -
sovereign-epistemic-commons-poc
Lightweight, replayable PoC for Sovereign Epistemic Commons: synthetic multi-agent memory-governance experiments on contamination, typed lanes, provenance-depth discount, and mixed exit/fork in observable-only agentic systems.
Python ★ 0 4mo agoExplain → -
rsi-yardstick-drift-poc
A lightweight, reproducible PoC series for studying recursive self-improvement under endogenous yardstick drift, using Gemma 3 local models to separate proxy-score drift, proxy gaming, delayed audit, and strict-confirmed improvement.
TeX ★ 0 4mo agoExplain → -
search-stability-lab
Theory-to-experiment lab for search stability in long-running agents under finite context, with exact simulator tests and lightweight mechanistic probe tasks.
Python ★ 0 4mo agoExplain → -
audit-closed-ai-scientist
Benchmark for statistically valid AI scientist systems, using audit-closed protocols, transparency logs, and sequential inference to prevent false discoveries in autonomous research agents.
Python ★ 0 4mo agoExplain → -
Proof-Carrying-Skills--PCS-Core-
Deterministic verifier and reference implementation for Proof-Carrying Skills (PCS-Core) and compute-saving inference reuse for AI/LLMs.
Python ★ 0 4mo agoExplain → -
observable-replay-lab
Observable-only no-meta epistemics lab: deterministic replay + reproducible audit logs, gate-based growth simulation, and identifiability/uncertainty benchmarks.
TeX ★ 0 4mo agoExplain → -
Oversight-Centered-Metrology-PoC
Lightweight proof-of-concept for oversight-centered metrology in coding agents: workflow-aware evaluation, interrupt channels, and claim-margin reporting beyond raw success scores.
Python ★ 0 4mo agoExplain → -
no-meta-drift-papers
Stage-based OSS packaging of no-meta + ontology-drift theory from four Zenodo preprints: paper-linked specs, AI-friendly manifests, and implementation stubs.
TeX ★ 0 4mo agoExplain → -
bottleneck-audit-toolkit
Offline, fail-closed verifier for JSONL telemetry event logs. Emits deterministic audit certificates + human summaries with explicit claims/non-claims for bottleneck and integrity review.
Python ★ 0 6mo agoExplain → -
Benevolent-Propagation-spec
No description.
★ 0 10mo agoExplain → -
UniverseModel
A Computational, Observer-Centric Universe Model
Python ★ 0 11mo agoExplain → -
AIconsciousness
conceptual AI swarm project
Python ★ 0 1y agoExplain → -
WisdomWeaver
Wisdom Weaver is an AI agent designed to help non-data science experts extract essential insights from limited data. This project provides a framework for robust analysis, especially for small datasets, leveraging the capabilities of the Gemini CLI.
Python ★ 0 1y agoExplain → -
ASI
ASI prototype
Python ★ 0 1y agoExplain →
No repos match these filters.