Pinned Loading
Repositories
- sentry-libraxis Public Forked from getsentry/self-hosted
Sentry, feature-complete and packaged up for low-volume deployments and proofs-of-concept
- vllm-swift Public Forked from TheTom/vllm-swift
vLLM Metal plugin powered by mlx-swift — high-performance LLM inference on Apple Silicon
- mlx-batch-server Public
High-performance MLX inference server for Apple Silicon — batch processing, Responses API, VLM support
- mlx-vlm Public Forked from Blaizzy/mlx-vlm
MLX-VLM is a package for inference and fine-tuning of Vision Language Models (VLMs) on your Mac using MLX.
- mlx-parallm Public
Batched KV caching for fast parallel inference on Apple Silicon via MLX - with streaming support
- mlx-omni-server Public Forked from madroidmaq/mlx-omni-server
MLX Omni Server is a local inference server powered by Apple's MLX framework, specifically designed for Apple Silicon (M-series) chips. It implements OpenAI-compatible API endpoints, enabling seamless integration with existing OpenAI SDK clients while leveraging the power of local ML inference.
People
This organization has no public members. You must be a member to see who’s a part of this organization.
Top languages
Loading…
Most used topics
Loading…