Skip to content
View gokay-ai's full-sized avatar
:octocat:
Open Source Maxxing
:octocat:
Open Source Maxxing

Highlights

  • Pro

Block or report gokay-ai

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
gokay-ai/README.md
Gökay Yılmaz — AI Engineer. Beyond the demo. Into the infrastructure.

I don't stop at using the AI stack. I build on it—and fix what breaks underneath.

I'm Gökay, an AI engineer working across LLM infrastructure, autonomous agents, and developer tools. From fine-tuning and RAG at Halkbank to upstream fixes in open-source AI, I work where model capability meets engineering reality.

Build the tool. Trace the failure. Ship the fix.

My projects · Open-source work · X · LinkedIn

Selected open-source contributions

Merged upstream

Hermes Agent · Nous Research
Agent infrastructure that recovers. Fixed interrupted-update recovery so pending gateway restarts complete on the next run—even when Git already reports up to date. Original contribution incorporated upstream with co-author credit.
View merged work ↗

OpenClaw
Recovery instead of failed turns. Fixed llama.cpp context-overflow detection to enable session compaction and retry. Also repaired plugin-catalog routing for direct navigation and refresh.
Agent recovery ↗ · Catalog routing ↗

Unsloth
Get past installation. Get to the model. Fixed Studio's Linux installation fallback when distro uv policy blocks automatic Python downloads.
View merged work ↗

Cherry Studio
Correctness at the data boundary. Fixed schema validation that broke knowledge-base listing when a database contained zero documents.
View merged work ↗

Ongoing contributions

Open pull requests across inference, training, and agent systems—not yet merged.

  • vLLM — KV-cache accounting for hidden-state extraction.
  • SGLang — Disaggregated-serving staging, request retraction, and vision preprocessing.
  • Hugging Face Accelerate — Avoiding double sharding of already-sharded data loaders.
  • OpenHands — Claude Code ACP skill integration and agent-server reliability.
  • LMDeploy — Safer HTTP weight updates and validation of untrusted serving inputs.
  • LibreChat — MCP instructions, agent generation parameters, and proxy support.
  • Jan — Custom OpenAI-compatible providers and llama.cpp runtime configuration.
  • Buzz — Bundled search-tool compatibility.

Explore the work →

Building

Sheep · Undo for AI coding agents

Agents move fast. Developers should keep control.

Turn-level checkpoints and worktree-scoped restores, written in Rust. Inspect the plan, rewind an agent's changes, and notify it of the restored state. Explicit previews. Scoped restores. Git metadata untouched.

Source · Releases · Architecture

Your agent finished. You should know.

Native completion notifications for macOS and Linux through Claude Code's Stop hook. Built for terminal-heavy workflows, including Ghostty, tmux, and Zellij.

Engineering focus

Python · Rust · TypeScript · PyTorch · Docker · Linux

LLM fine-tuning and RAG experience at Halkbank. Studying Artificial Intelligence at the University of Birmingham, graduating in 2027. Building toward agents with better memory, measurable behavior, and recoverable workflows.


Building serious AI infrastructure or developer tools? Let's talk.
X / @gokayai · LinkedIn

Pinned Loading

  1. sheep sheep Public

    Undo for AI coding agents. Every agent turn becomes a restorable checkpoint.

    Rust 6

  2. openclaw/openclaw openclaw/openclaw Public

    The AI that really does things. Any OS. Any Platform. The lobster way. 🦞

    TypeScript 390k 81.9k

  3. NousResearch/hermes-agent NousResearch/hermes-agent Public

    The agent that grows with you

    Python 246k 51.4k

  4. vllm-project/vllm vllm-project/vllm Public

    A high-throughput and memory-efficient inference and serving engine for LLMs

    Python 91.9k 22.3k

  5. unslothai/unsloth unslothai/unsloth Public

    Local UI to run and train LLMs and diffusion models. Supports GGUF, MLX, Qwen3.8, DeepSeek-V4, MiniMax-H3, Gemma 4, FLUX and more.

    Python 76.2k 6.9k

  6. CherryHQ/cherry-studio CherryHQ/cherry-studio Public

    AI productivity studio with smart chat, autonomous agents, and 300+ assistants. Unified access to frontier LLMs

    TypeScript 51.9k 5k