Skip to content
View eduardopessin's full-sized avatar

Block or report eduardopessin

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
eduardopessin/README.md

LLM infrastructure, mostly around LiteLLM. I like problems where the answer comes from measuring the wire rather than reading the docs, and most of what I publish started as something that was broken in my own cluster.

What I'm working on

litellm-mysubs serves Claude Max, ChatGPT Plus and Google Antigravity subscriptions as ordinary LiteLLM models, so any OpenAI-compatible client can use them. One line in callbacks: installs it. On PyPI.

crowdcompute is a demand-validation MVP for community-funded European AI compute: 300 people, one 8xH200 cluster.

cobol4i migrates IBM i ILE COBOL to Java through deterministic gates first, and only then lets a model touch what is left.

Upstream

ggml-org/llama.cpp#24710 (merged), making tensor-split regex patterns static so they are not recompiled per tensor.

BerriAI/litellm#42467, where a provider reporting a model name the price map does not carry silently bills the turn at zero. The issue has the measurements, including the two wrong theories I had to walk back first.

I also filed BerriAI/litellm#42172, where an Anthropic OAuth token from a Claude Code client reaches third-party deployments and replaces their configured key.

How I work

Measure before claiming. Most of my issue reports and commit messages carry the numbers that led to the conclusion, and when the conclusion turns out wrong I correct it in the same thread rather than quietly moving on. Tests are there to fail when the code breaks, not to make coverage look good.

Pinned Loading

  1. tokengateway tokengateway Public

    OAuth token manager, subscription quota dashboard and native wire-protocol gateway. The LiteLLM plugin part now lives in litellm-mysubs.

    Python 4 2

  2. litellm-mysubs litellm-mysubs Public

    Serve your Claude Max, ChatGPT Plus and Google Antigravity subscriptions as LiteLLM models. One command to install.

    Python 2 1

  3. cobol4i cobol4i Public

    COBOL for i — migrate IBM i (AS/400) ILE COBOL to Java through deterministic gates, then let an LLM refactor it and prove behaviour did not change.

    COBOL

  4. crowdcompute crowdcompute Public

    Community-funded European AI compute: 300 people, one 8xH200 cluster. Demand-validation MVP live at crowdcompute.eu (Astro, Cloudflare Pages, D1, Turnstile).

    Astro

  5. consumer-multigpu-inference consumer-multigpu-inference Public

    PCIe BAR1 P2P on consumer Blackwell: driver patch, operational tooling and benchmark methodology for tensor-parallel inference on 4x RTX 5060 Ti

    Python