Skip to content
@WorldFlowAI

WorldFlowAI

Popular repositories Loading

  1. everything-claude-code everything-claude-code Public

    Claude Code toolkit - agents, commands, skills, rules, and hooks for productive AI-assisted development

    JavaScript 2.8k 432

  2. synapse-memory synapse-memory Public

    MCP server for Claude Code — persistent session memory with zero infrastructure

    TypeScript 5 2

  3. semblend semblend Public

    Semantic KV cache reuse for LLM inference engines (vLLM, SGLang)

    Python 5 2

  4. litellm-to-synapse litellm-to-synapse Public

    Migration tool for moving from LiteLLM to Synapse — converts LiteLLM YAML config to Synapse API payloads

    Python 1

  5. vllm vllm Public

    Forked from vllm-project/vllm

    A high-throughput and memory-efficient inference and serving engine for LLMs

    Python 1

  6. LMCache LMCache Public

    Forked from LMCache/LMCache

    Supercharge Your LLM with the Fastest KV Cache Layer

    Python 1

Repositories

Showing 10 of 13 repositories
  • sglang Public Forked from sgl-project/sglang

    SGLang is a high-performance serving framework for large language models and multimodal models.

    WorldFlowAI/sglang's past year of commit activity
    Python 1 Apache-2.0 8,822 0 1 Updated Sep 8, 2026
  • sembench Public

    SemBench: benchmark suite for semantic KV cache reuse

    WorldFlowAI/sembench's past year of commit activity
    Python 1 0 0 0 Updated Sep 3, 2026
  • semblend Public

    Semantic KV cache reuse for LLM inference engines (vLLM, SGLang)

    WorldFlowAI/semblend's past year of commit activity
    Python 5 2 1 0 Updated Sep 4, 2026
  • semblend-vllm-connector Public

    Out-of-tree vLLM KVConnector for SemBlend semantic KV donor discovery

    WorldFlowAI/semblend-vllm-connector's past year of commit activity
    Python 1 Apache-2.0 1 0 4 Updated Sep 3, 2026
  • vllm Public Forked from vllm-project/vllm

    A high-throughput and memory-efficient inference and serving engine for LLMs

    WorldFlowAI/vllm's past year of commit activity
    Python 1 Apache-2.0 22,284 0 0 Updated Sep 4, 2026
  • peregrine Public

    FFmpeg for AI inference: a pure-C inference engine with hand-written assembly (no intrinsics), x86-64 + ARM, runtime CPU dispatch.

    WorldFlowAI/peregrine's past year of commit activity
    C 1 BSD-2-Clause 0 0 1 Updated Jun 21, 2026
  • dynamo Public Forked from ai-dynamo/dynamo

    A Datacenter Scale Distributed Inference Serving Framework

    WorldFlowAI/dynamo's past year of commit activity
    Rust 1 1,588 0 0 Updated Jun 18, 2026
  • dynamo-fork Public Forked from ai-dynamo/dynamo

    A Datacenter Scale Distributed Inference Serving Framework

    WorldFlowAI/dynamo-fork's past year of commit activity
    Rust 1 1,588 0 0 Updated Mar 28, 2026
  • LMCache Public Forked from LMCache/LMCache

    Supercharge Your LLM with the Fastest KV Cache Layer

    WorldFlowAI/LMCache's past year of commit activity
    Python 1 Apache-2.0 1,895 0 0 Updated Mar 18, 2026
  • litellm-to-synapse Public

    Migration tool for moving from LiteLLM to Synapse — converts LiteLLM YAML config to Synapse API payloads

    WorldFlowAI/litellm-to-synapse's past year of commit activity
    Python 1 0 0 0 Updated Mar 13, 2026

People

This organization has no public members. You must be a member to see who’s a part of this organization.

Most used topics

Loading…