Skip to content
View LobsterQBA's full-sized avatar
🎯
Focusing
🎯
Focusing

Block or report LobsterQBA

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
LobsterQBA/README.md

Hi, I'm Leo Zhao

Data scientist in Seattle building AI tools, agent systems, and decision products. I care about systems that are useful, inspectable, and honest about where people still need to decide.

leozhao.me · Interested in agent reliability, useful evaluation, and human-in-the-loop product design.

Selected projects

Project What it is
Loop Agent · demo A tool-using agent loop with persistent traces, integrity checks, and bounded tool calls, arguments, and results. Runs locally without an API key.
Tab Tidy · Chrome Web Store Chrome extension that proposes tab groups, takes plain-English refinements, and previews every change before applying it.
Point2Prompt · install Bookmarklet that turns a click on any UI element into a structured change brief for Claude Code, Codex, or Cursor.
Trackpad Canvas · site Native macOS diagramming app that draws from raw trackpad touches and snaps sketches into connected architecture diagrams.
SplitTaste · demo Repairs shared-account streaming recommendations with two user questions; reproducible MovieLens 32M evaluation with honest metric gates.
Where to Sit · demo 3D IMAX seat-view simulator for 20 venues across Seattle, NYC, and the Bay Area.

Open-source contributions

  • Microsoft Agent Framework — surfaced A2A preview consent URLs and added regression coverage.
  • Strands Harness SDK — stops retry backoff promptly when a TypeScript run is cancelled.
  • OpenMed — adds strict, versioned parsing for agent run summaries.
  • DeepEval — normalizes verbose judge verdicts so ambiguous outputs cannot silently become passing scores.
  • OpenHarness — prevents disabled tools from leaking into model guidance.

Pinned Loading

  1. loop-agent loop-agent Public

    A Python app that shows how an agent calls tools, saves results, and recalls them after a restart.

    Python 44 5

  2. microsoft/agent-framework microsoft/agent-framework Public

    A framework for building, orchestrating and deploying AI agents and multi-agent workflows with support for Python and .NET.

    Python 13.7k 2.4k

  3. HKUDS/OpenHarness HKUDS/OpenHarness Public

    "OpenHarness: Open Agent Harness with a Built-in Personal Agent--Ohmo!"

    Python 15.8k 2.6k

  4. strands-agents/harness-sdk strands-agents/harness-sdk Public

    Build an agent harness and control it end-to-end. Open-source SDK for production AI agents in Python & TypeScript - any model, any cloud.

    Python 7.4k 1.2k

  5. mlflow/mlflow mlflow/mlflow Public

    The open source AI engineering platform for agents, LLMs, and ML models. MLflow enables teams of all sizes to debug, evaluate, monitor, and optimize production-quality AI applications while control…

    Python 28.1k 6.3k

  6. confident-ai/deepeval confident-ai/deepeval Public

    The LLM Evaluation Framework

    Python 18.4k 2k