AI Quality Engineer — LLM evaluation, safety & red teaming. 9+ yrs in QA.
Pinned Loading
-
agentic-api-qa
agentic-api-qa PublicMulti-agent API QA harness on LangGraph — Explorer and Adversary LLM agents probe a live API under a deterministic safety governor, with an LLM judge, bounded autonomy, prompt-injection-resistant e…
Python
-
llm-eval-nanotech
llm-eval-nanotech PublicLLM evaluation suite for the nanotech.icu AI chatbot — DeepEval metrics, a custom LLM judge, nightly CI gates, and Pytest/Playwright E2E and API coverage.
Python
-
alexpavsky
alexpavsky PublicSource of alexpavsky.com — a self-hosted AI assistant with a free-model failover pool, RAG, a voice agent, and a live tech feed.
Python
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.




