Skip to content
View ryangu00's full-sized avatar
🎯
Focusing
🎯
Focusing

Highlights

  • Pro

Block or report ryangu00

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
ryangu00/README.md

Ryan Gu

I run AI models and coding agents on my own hardware and write down what I measure: what worked, what failed, and how to check it yourself.

Open-source contributions

My projects

  • axiom: checks an AI agent's "done" claims against evidence declared before the work.
  • ryanai-evalbank: a method for scoring a model on your own work without publishing the questions.
  • ryanai-lab-cookbooks: an index of 54 RyanAI Lab cookbooks on serving, training, evaluation and coding agents.

Pinned Loading

  1. axiom axiom Public

    Prove every "done." — verifies AI agent completion claims against evidence declared before the work. Observe-mode by default, zero-config, four runtimes.

    Python 4 1

  2. ryanai-evalbank ryanai-evalbank Public

    evalbank: a method and harness for scoring a model on your own work, without publishing the questions

    Python

  3. ryanai-lab-cookbooks ryanai-lab-cookbooks Public

    Index of the RyanAI Lab cookbooks: five reader paths, nine sections, 54 cookbooks with a current/rollback/historical status

  4. training-reranker training-reranker Public

    Self-hosted reranker and the missing-scoring-head GGUF trap that fakes a weak model

    Python

  5. agentic-rag-from-scratch agentic-rag-from-scratch Public

    A six-stage agentic RAG loop in ~200 lines of pure-stdlib Python — gap critic, adversarial verification, cited synthesis. Retrieval is a loop, not a lookup.

    Python

  6. training-embedding-finetune training-embedding-finetune Public

    Fine-tune your own embedding model: pipeline scripts, Ollama deployment, and three production incidents

    Shell