ray-llm
by Community
RayLLM - LLMs on Ray (Archived). Read README for more info.
OSS
ray-llm
Added 1 June 2026
Overview
RayLLM is a community archive repository offering tools for running large language models on the Ray distributed compute framework. The README provides specific setup and usage details, but the project is no longer actively maintained.
Best for
Best for
Developers already using Ray who need legacy code or patterns for running LLMs at scale.
Use cases
- Deploying open-source LLMs on a Ray cluster
- Scaling LLM inference across multiple nodes
- Testing Ray-based orchestration for model serving
Notes
RayLLM is a community archive repository offering tools for running large language models on the Ray distributed compute framework. The README provides specific setup and usage details, but the project is no longer actively maintained.
1,267 stars on GitHub. Last updated 2025-03-13.
Use cases
- Deploying open-source LLMs on a Ray cluster
- Scaling LLM inference across multiple nodes
- Testing Ray-based orchestration for model serving
Pros
- Leverages Ray’s distributed computing for large models
- Open source with a public archive for reference
- Straightforward integration with Ray ecosystem
Cons
- Archived and not actively maintained or updated
- Limited community support beyond existing documentation
- May lack compatibility with newer Ray versions or LLM frameworks
Indexed from awesome-llmops and enriched against its public facts.
Pros
- Leverages Ray's distributed computing for large models
- Open source with a public archive for reference
- Straightforward integration with Ray ecosystem
Cons
- Archived and not actively maintained or updated
- Limited community support beyond existing documentation
- May lack compatibility with newer Ray versions or LLM frameworks
Open-source & AI alternatives
Swap-in tools that solve the same job. Weigh the trade-offs before you commit.
vLLM
Community
A high-throughput and memory-efficient inference and serving engine for LLMs
FastChat
Community
An open platform for training, serving, and evaluating large language models. Release repo for Vicuna and Chatbot Arena.
SGLang
Community
SGLang is a high-performance serving framework for large language models and multimodal models.
Get the free Developer’s Field Guide
A 27-page field guide to the AI coding workflow with Claude. Claude Code, MCP servers, the prompt patterns that work, and what to delegate. Free.
Enter your work email. We send it straight over, plus a few short notes worth knowing. Unsubscribe any time.
