Skip to content
View f0909172434's full-sized avatar

Highlights

  • Pro

Block or report f0909172434

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
f0909172434/README.md

Chih-Kai Wang banner

English · 繁體中文 · 简体中文

Chih-Kai Wang (王治凱)

Building verifiable agent systems, sandboxed execution runtimes, and formal mathematics tools. 致力於打造可驗證的自主 Agent 架構、內核沙盒執行環境與形式化數學證明系統。

Mathematics Division, Department of Mathematics and Information Education, National Taipei University of Education (NTUE), expected 2028. Based in Taipei, Taiwan.
國立臺北教育大學 數學暨資訊教育學系 數學組(預計 2028 年畢業)。

📑 Curriculum Vitae (CV) · 🐙 GitHub · ✉️ Email


🌟 Featured Research & Engineering Systems

🤖 Autonomous Agent Architectures & Sandboxed Runtimes

  • DSH Architecture Lab
    Production-grade autonomous agent research laboratory for DeepSeek Harness.
    Features kernel-level Lima Linux VM isolation, micro-cent token broker metering, immutable cryptographic research protocols (review-protocol.json), and factorial empirical trials across multiple memory and planning recipes (A/B/C/D). Verified 100% test pass rate with DeepSeek-V4.1-Flash.
  • RuleShift
    Deterministic local testbed evaluating agent memory adaptation under dynamic rule changes.
    Encompasses 800 paired tasks with automated state verification and replayable evidence. Empirically demonstrated that foundational retrieval delivers top-tier 95.6% success rates with superior cost efficiency.
  • DSH Second Agent Kit
    macOS defense-in-depth security and memory isolation suite for DeepSeek Harness.
    Combines macOS Seatbelt kernel profiles (network restrictions, local socket channels), interactive circuit breakers (anti-deadlock protection), and AsyncLocalStorage-based project memory isolation.
  • MiniHarness
    Comprehensive 8-step AI Agent Harness engineering curriculum and workshop in Python.
    Provides an offline-first learning path from tool dispatch and agent loops to memory palaces and production-grade architectures, accompanied by a complete Traditional Chinese curriculum.
  • Verified Search Plugin
    Source-verifiable, hallucination-resistant retrieval workflow for DeepSeek Harness.
    Backed by 250+ unit tests and 42 frozen benchmark corpora, featuring structured JSON extraction and explicit evidence-gap tracking (unresolved).

📐 Formal Mathematics & Automated Theorem Proving

  • ProofWeave Core
    Structured mathematical claim verification platform powered by Lean 4.
    Transforms natural and structured mathematical arguments into inspectable certification runs, rigorously separating formal proof validity (Lean 4 kernel checked) from semantic statement alignment.
  • SAIR Proof Press
    Equational implication solver companion with heuristic search and Lean 4 certificate generation.
    Provides public companion datasets, immutable benchmark artifacts, released-input evaluation suites, and an accompanying academic paper.
  • Finite Witness
    WebMCP-integrated finite-graph counterexample search tool.
    Features bounded exhaustive search, client-side data sovereignty, inspectable certificates, and an independent Python replay verifier.

🛡️ Engineering Quality, CI & Reproducibility

  • HonestCI
    Developer-first GitHub Action and CLI that makes green CI truly meaningful.
    Wraps test execution, verifies fresh and uncorrupted JUnit XML evidence, guards against test-count drops relative to trusted default branches, and detects false-green CI anti-patterns (v1.0.4 on npm & GitHub Marketplace).
  • RigorGraph
    Local-first claim-evidence DAGs and deterministic audit reports.
    Enforces four-eyes review principles (separation of author and independent reviewer duties), detects DAG cycles, computes SHA-256 byte fingerprints, and outputs portable offline HTML audit dossiers.

🔬 Interactive AI Labs & Specialized Tools

  • TokenScope
    Interactive bilingual browser laboratory for LLM causal attention, token sampling, and BPE.
    Features inspectable $5 \times 5$ self-attention matrices, temperature/top-k/top-p probability distribution controls, and a fully reversible Unicode BPE tokenizer.
  • Charlie Alpha 4B
    Trilingual statistical procedure-selection model optimized for Apple Silicon MLX.
    Provides reproducible statistical decision assistance with built-in cautious clarification fallbacks (needs_clarification) for ambiguous problem formulations.
  • DeepSeek Girl Pets
    16-direction animated companion plugin for DeepSeek Harness & Codex Desktop.
    Zero DOM pollution, native lifecycle hook synchronization, and 100% offline privacy preservation.

🛠️ Technical Competencies

Domain Technologies & Frameworks
Languages Python, TypeScript, JavaScript (Node.js 20+), Lean 4, SQL, Bash/Zsh
Agent & Harness Engineering DeepSeek Harness, Model Context Protocol (MCP), WebMCP, Tool Calling Loops, Memory Palaces
Sandboxing & Isolation Lima Linux VM, Docker, macOS Seatbelt (sandbox-exec), Process Isolation, Circuit Breakers
Formal Methods & Math Lean 4 & Mathlib, Proof Assistants, Equational Logic, Finite Graph Search, Abstract Algebra
AI & Machine Learning Apple Silicon MLX, PyTorch, Causal Self-Attention, BPE Tokenization, Local LLM Quantization
Quality & Infrastructure GitHub Actions, HonestCI, JUnit XML Standards, Vitest / Pytest, Prettier / ESLint / Ruff

💡 Engineering Philosophy

  1. Keep Evidence Attached to Claims: Claims without verifiable computational evidence remain conjectures. Every conclusion must be grounded in replayable artifacts.
  2. Defense-in-Depth & Fail-Closed Design: Sandbox boundaries, Token budgets, and execution timeouts must fail closed, protecting system integrity and user resources.
  3. Determinism & Reproducibility: Research tools and benchmarks should operate 100% offline with pinned seeds and cryptographic hashes, eliminating dependence on external network states.
  4. Constructive Open-Source Craftsmanship: Deliver robust, well-documented, and human-friendly software that empowers developers, students, and researchers alike.

Designed with precision in Taipei · Open to software engineering & AI research internships.

Pinned Loading

  1. honest-ci honest-ci Public

    Make green CI mean the tests you expected actually ran.

    TypeScript 3

  2. finite-witness-webmcp finite-witness-webmcp Public

    Finite-graph counterexample search with inspectable certificates, independent Python replay, and eight shared WebMCP tools.

    JavaScript

  3. rigorgraph rigorgraph Public

    Local-first claim-evidence graphs and deterministic audit reports for AI-assisted research.

    Python 1

  4. proofweave-math-lab proofweave-math-lab Public

    Experimental Python and Lean checker for structured mathematical claims, with inspectable certificates and explicit semantic-alignment limits.

    Python 2

  5. sair-stage2-proof-press sair-stage2-proof-press Public

    Public companion to Lean-checked equational implication solvers: frozen artifacts, released-input evaluations, and an English research paper.

    TypeScript 1

  6. tokenscope tokenscope Public

    An interactive bilingual lab for causal attention, token sampling, and BPE. Inspect the numbers, change the controls, export the experiment.

    TypeScript