Releases: AtomicBot-ai/dflash
Release list
DFlash MLX Server macOS ARM64 (fe12300)
DFlash MLX Server — macOS ARM64
Built from main branch at commit fe12300.
What's included
mlx-server— standalone OpenAI-compatible inference server- Standard MLX inference (mlx-lm)
- DFlash speculative decoding acceleration
Usage
tar -xzf dflash-mlx-server-macos-arm64.tar.gz
# Standard MLX inference
./mlx-server --model /path/to/model --port 8080
# DFlash accelerated inference
./mlx-server --model /path/to/model --draft-model /path/to/draft --port 8080For Atomic Chat integration
This binary is automatically downloaded during the Atomic Chat build process.
DFlash MLX Server macOS ARM64 (949c1c4)
DFlash MLX Server — macOS ARM64
Built from main branch at commit 949c1c4.
What's included
mlx-server— standalone OpenAI-compatible inference server- Standard MLX inference (mlx-lm)
- DFlash speculative decoding acceleration
Usage
tar -xzf dflash-mlx-server-macos-arm64.tar.gz
# Standard MLX inference
./mlx-server --model /path/to/model --port 8080
# DFlash accelerated inference
./mlx-server --model /path/to/model --draft-model /path/to/draft --port 8080For Atomic Chat integration
This binary is automatically downloaded during the Atomic Chat build process.
DFlash MLX Server macOS ARM64 (7ac8b49)
DFlash MLX Server — macOS ARM64
Built from main branch at commit 7ac8b49.
What's included
mlx-server— standalone OpenAI-compatible inference server- Standard MLX inference (mlx-lm)
- DFlash speculative decoding acceleration
Usage
tar -xzf dflash-mlx-server-macos-arm64.tar.gz
# Standard MLX inference
./mlx-server --model /path/to/model --port 8080
# DFlash accelerated inference
./mlx-server --model /path/to/model --draft-model /path/to/draft --port 8080For Atomic Chat integration
This binary is automatically downloaded during the Atomic Chat build process.
DFlash MLX Server macOS ARM64 (61b57ba)
DFlash MLX Server — macOS ARM64
Built from main branch at commit 61b57ba.
What's included
mlx-server— standalone OpenAI-compatible inference server- Standard MLX inference (mlx-lm)
- DFlash speculative decoding acceleration
Usage
tar -xzf dflash-mlx-server-macos-arm64.tar.gz
# Standard MLX inference
./mlx-server --model /path/to/model --port 8080
# DFlash accelerated inference
./mlx-server --model /path/to/model --draft-model /path/to/draft --port 8080For Atomic Chat integration
This binary is automatically downloaded during the Atomic Chat build process.
DFlash MLX Server macOS ARM64 (1936b7a)
DFlash MLX Server — macOS ARM64
Built from main branch at commit 1936b7a.
What's included
mlx-server— standalone OpenAI-compatible inference server- Standard MLX inference (mlx-lm)
- DFlash speculative decoding acceleration
Usage
tar -xzf dflash-mlx-server-macos-arm64.tar.gz
# Standard MLX inference
./mlx-server --model /path/to/model --port 8080
# DFlash accelerated inference
./mlx-server --model /path/to/model --draft-model /path/to/draft --port 8080For Atomic Chat integration
This binary is automatically downloaded during the Atomic Chat build process.
DFlash MLX Server macOS ARM64 (2c70d44)
DFlash MLX Server — macOS ARM64
Built from main branch at commit 2c70d44.
What's included
mlx-server— standalone OpenAI-compatible inference server- Standard MLX inference (mlx-lm)
- DFlash speculative decoding acceleration
Usage
tar -xzf dflash-mlx-server-macos-arm64.tar.gz
# Standard MLX inference
./mlx-server --model /path/to/model --port 8080
# DFlash accelerated inference
./mlx-server --model /path/to/model --draft-model /path/to/draft --port 8080For Atomic Chat integration
This binary is automatically downloaded during the Atomic Chat build process.
DFlash MLX Server macOS ARM64 (1ed2330)
DFlash MLX Server — macOS ARM64
Built from main branch at commit 1ed2330.
What's included
mlx-server— standalone OpenAI-compatible inference server- Standard MLX inference (mlx-lm)
- DFlash speculative decoding acceleration
Usage
tar -xzf dflash-mlx-server-macos-arm64.tar.gz
# Standard MLX inference
./mlx-server --model /path/to/model --port 8080
# DFlash accelerated inference
./mlx-server --model /path/to/model --draft-model /path/to/draft --port 8080For Atomic Chat integration
This binary is automatically downloaded during the Atomic Chat build process.
DFlash MLX Server macOS ARM64 (89d59f0)
DFlash MLX Server — macOS ARM64
Built from main branch at commit 89d59f0.
What's included
mlx-server— standalone OpenAI-compatible inference server- Standard MLX inference (mlx-lm)
- DFlash speculative decoding acceleration
Usage
tar -xzf dflash-mlx-server-macos-arm64.tar.gz
# Standard MLX inference
./mlx-server --model /path/to/model --port 8080
# DFlash accelerated inference
./mlx-server --model /path/to/model --draft-model /path/to/draft --port 8080For Atomic Chat integration
This binary is automatically downloaded during the Atomic Chat build process.
DFlash MLX Server macOS ARM64 (f486034)
DFlash MLX Server — macOS ARM64
Built from main branch at commit f486034.
What's included
mlx-server— standalone OpenAI-compatible inference server- Standard MLX inference (mlx-lm)
- DFlash speculative decoding acceleration
Usage
tar -xzf dflash-mlx-server-macos-arm64.tar.gz
# Standard MLX inference
./mlx-server --model /path/to/model --port 8080
# DFlash accelerated inference
./mlx-server --model /path/to/model --draft-model /path/to/draft --port 8080For Atomic Chat integration
This binary is automatically downloaded during the Atomic Chat build process.
DFlash MLX Server macOS ARM64 (b742c32)
DFlash MLX Server — macOS ARM64
Built from main branch at commit b742c32.
What's included
mlx-server— standalone OpenAI-compatible inference server- Standard MLX inference (mlx-lm)
- DFlash speculative decoding acceleration
Usage
tar -xzf dflash-mlx-server-macos-arm64.tar.gz
# Standard MLX inference
./mlx-server --model /path/to/model --port 8080
# DFlash accelerated inference
./mlx-server --model /path/to/model --draft-model /path/to/draft --port 8080For Atomic Chat integration
This binary is automatically downloaded during the Atomic Chat build process.