Skip to content
View sujee's full-sized avatar

Highlights

  • Pro

Block or report sujee

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
sujee/README.md

Hi, I'm Sujee πŸ‘‹

I'm an AI Developer Advocate at Nebius, helping developers build with open models through technical content, practical examples, guides, demos, and hands-on workshops.

I bring 25+ years of software-engineering experience across distributed systems, data engineering, machine learning, and cloud infrastructure.

πŸ”— Connect with me

🌐 sujee.dev β€’ πŸ™ GitHub β€’ πŸ’Ό LinkedIn β€’ 🐦 X β€’ πŸ¦‹ Bluesky β€’ πŸŽ₯ YouTube β€’ πŸ’¬ Discord: @sujee.dev

πŸ› οΈ Projects

Quick, real-time LLM endpoint benchmarks from your browser. Measures speed, throughput, latency, accuracy, token usage, and estimated cost, with no setup required. Try it live β†’

Practical benchmarks, performance tests, and hands-on evaluations for open models, including a visual explorer for models on Nebius Token Factory. Explore models β†’

Open-source, reproducible visual explainers for AI and software concepts. Watch them, reproduce them, remix them. Watch on YouTube β†’

A live visual arena where two LLMs battle at Snake, a fun way to compare model latency, reliability, and decision-making across OpenAI-compatible APIs. Try it live β†’

A full-stack, open-source RAG chatbot that answers questions about your website using open models, document processing, embeddings, and vector databases.

Practical examples using Data Prep Kit, showcasing document processing, data preparation, Docling, Milvus, and open-source RAG systems.

πŸ”­ What I'm working on and exploring

  • πŸ’» Coding with open models β€” how open models perform in real coding workflows with Claude Code, Codex, OpenCode, Cline, and Cursor
  • πŸ§ͺ Evals β€” practical benchmarks for models, endpoints, and model + agent-harness combinations Β· Practical LLM Evals
  • ⚑ Inference β€” LLM performance, latency, throughput, and cost Β· ZebraBench, LLM Snake Arena
  • πŸ€– AI agents β€” agentic workflows and tool use
  • 🧹 Data prep β€” document processing and data preparation for AI Β· Data Prep Kit Examples
  • 🦾 Physical AI β€” robotics and embodied AI

🎀 I speak and run hands-on workshops on these topics, and make open-source Visual Explainers for AI concepts. Get in touch.

🧰 Technology stack

Models and inference: Open models, Ollama, Nebius Token Factory, hosted and self-hosted inference

AI and agent frameworks: LangChain, Deep Agents, Tavily, OpenWiki, Docling

Languages and infrastructure: Python, Docker, Kubernetes, cloud infrastructure

Pinned Loading

  1. llm-snake-arena llm-snake-arena Public

    🐍 A live visual arena where LLMs compete in Snake - compare model latency, reliability and decision-making across OpenAI-compatible APIs.

    JavaScript

  2. practical-llm-evals practical-llm-evals Public

    Practical benchmarks, performance tests, and experiential evaluations for open models.

    JavaScript

  3. visual-stories visual-stories Public

    Open-source, reproducible visual explainers for AI and software concepts. Watch them. Reproduce them. Remix them.

    Python

  4. zebrabench zebrabench Public

    Quick, real-time LLM endpoint benchmarks from your browser.

    JavaScript