Skip to content
@sybil-solutions

Sybil

A place for all things AI
Sybil Solutions

Sybil Solutions

A place for all things AI.

Website X Email Repos


Sybil Solutions builds local-first tooling for running, steering, and using self-hosted LLM backends. Everything here assumes you own the box it runs on.

Repositories

Repository Language Stars Description
local-studio TypeScript stars Control panel for VLLM, SGLang, llama.cpp, ExLlamaV3
codex-shim Python stars Local Responses-API shim exposing Factory BYOK models (and optional ChatGPT GPT-5.5 passthrough) to Codex Desktop

Focus areas

  • Local model lifecycle — install, launch, and supervise inference engines on your own hardware.
  • OpenAI-compatible proxies — point existing tools at local endpoints without changing call sites.
  • Agent runtimes — run coding agents against local or remote controllers from one surface.

Links

Profile README lives in sybil-solutions/.github. Edit profile/README.md to update this page.

Popular repositories Loading

  1. local-studio local-studio Public

    Control panel for VLLM, Sglang, llama.cpp, exllamav3

    TypeScript 1.8k 168

  2. ai-data-extraction ai-data-extraction Public

    extract all your personal data history from cursor, codex, claude-code, windsurf, and trae

    Python 1.3k 113

  3. codex-shim codex-shim Public

    Local Responses-API shim that exposes Factory BYOK models (and optional ChatGPT GPT-5.5 passthrough) to Codex Desktop.

    Python 1.1k 103

  4. glm-flash-lite glm-flash-lite Public

    GLM-5.3-Flash EXL3 on one 24 GB RTX 3090 + DDR4: elastic GPU expert cache, zero-copy experts, AVX2 CPU tier, OpenAI API

    Python 290 27

  5. local-ai-registry local-ai-registry Public

    Local AI registry: one validated recipe per machine, with the evidence attached

    HTML 246 53

  6. dsv41-flash-offload dsv41-flash-offload Public

    DeepSeek-V4.1-Flash EXL3 on one 24 GB RTX 3090 + DDR4 + NVMe: staged-DMA prefill, AVX2 CPU expert tier, elastic VRAM expert cache, Engram on disk, OpenAI API

    Python 138 22

Repositories

Showing 10 of 22 repositories
  • omarchy-local-ai Public

    Local AI for Omarchy: the model validated for your GPU, one button on the bar. Start serves it, any coding agent opens on it, one click shares it on your tailnet.

    sybil-solutions/omarchy-local-ai's past year of commit activity
    Shell 135 MIT 24 4 5 Updated Oct 11, 2026
  • local-ai-usage Public

    Install counter for the Local AI Omarchy plugin (Cloudflare Worker + D1)

    sybil-solutions/local-ai-usage's past year of commit activity
    JavaScript 0 0 0 0 Updated Oct 11, 2026
  • moetier Public

    Minimal data-first standard for MoE expert placement and scheduling across VRAM, RAM and NVMe (records + 500-line core + simulator)

    sybil-solutions/moetier's past year of commit activity
    Python 5 MIT 2 0 0 Updated Oct 10, 2026
  • local-ai-registry Public

    Local AI registry: one validated recipe per machine, with the evidence attached

    sybil-solutions/local-ai-registry's past year of commit activity
    HTML 246 MIT 53 3 26 Updated Oct 9, 2026
  • glm-flash-lite Public

    GLM-5.3-Flash EXL3 on one 24 GB RTX 3090 + DDR4: elastic GPU expert cache, zero-copy experts, AVX2 CPU tier, OpenAI API

    sybil-solutions/glm-flash-lite's past year of commit activity
    Python 290 MIT 27 3 0 Updated Oct 9, 2026
  • local-ai-images Public

    Attested container images pinned by the local-ai registry

    sybil-solutions/local-ai-images's past year of commit activity
    Python 8 3 1 3 Updated Oct 9, 2026
  • qwen38-flash-next-b70-offload Public

    Qwen3.8-Flash-Next on one Intel Arc Pro B70 with experts on NVMe and 32 GB host RAM: SGLang + exl3xpu tier, staged NVMe prefill, SYCL sparse attention, measured results

    sybil-solutions/qwen38-flash-next-b70-offload's past year of commit activity
    Python 21 MIT 1 1 0 Updated Oct 8, 2026
  • ai-data-extraction Public

    extract all your personal data history from cursor, codex, claude-code, windsurf, and trae

    sybil-solutions/ai-data-extraction's past year of commit activity
    Python 1,345 113 2 5 Updated Oct 7, 2026
  • dsv41-flash-offload Public

    DeepSeek-V4.1-Flash EXL3 on one 24 GB RTX 3090 + DDR4 + NVMe: staged-DMA prefill, AVX2 CPU expert tier, elastic VRAM expert cache, Engram on disk, OpenAI API

    sybil-solutions/dsv41-flash-offload's past year of commit activity
    Python 138 MIT 22 1 0 Updated Oct 7, 2026
  • trellis-serve Public

    EXL3 (ExLlamaV3 trellis quantization) in stock SGLang and vLLM on RTX 3090 and Intel Arc B70, bit-exact with ExLlamaV3

    sybil-solutions/trellis-serve's past year of commit activity
    Python 2 MIT 3 0 4 Updated Oct 7, 2026

People

This organization has no public members. You must be a member to see who’s a part of this organization.

Most used topics

Loading…