Skip to content

Releases: AtomicBot-ai/dflash

DFlash MLX Server macOS ARM64 (fe12300)

Choose a tag to compare

DFlash MLX Server — macOS ARM64

Built from main branch at commit fe12300.

What's included

  • mlx-server — standalone OpenAI-compatible inference server
  • Standard MLX inference (mlx-lm)
  • DFlash speculative decoding acceleration

Usage

tar -xzf dflash-mlx-server-macos-arm64.tar.gz

# Standard MLX inference
./mlx-server --model /path/to/model --port 8080

# DFlash accelerated inference
./mlx-server --model /path/to/model --draft-model /path/to/draft --port 8080

For Atomic Chat integration

This binary is automatically downloaded during the Atomic Chat build process.

DFlash MLX Server macOS ARM64 (949c1c4)

Choose a tag to compare

DFlash MLX Server — macOS ARM64

Built from main branch at commit 949c1c4.

What's included

  • mlx-server — standalone OpenAI-compatible inference server
  • Standard MLX inference (mlx-lm)
  • DFlash speculative decoding acceleration

Usage

tar -xzf dflash-mlx-server-macos-arm64.tar.gz

# Standard MLX inference
./mlx-server --model /path/to/model --port 8080

# DFlash accelerated inference
./mlx-server --model /path/to/model --draft-model /path/to/draft --port 8080

For Atomic Chat integration

This binary is automatically downloaded during the Atomic Chat build process.

DFlash MLX Server macOS ARM64 (7ac8b49)

Choose a tag to compare

DFlash MLX Server — macOS ARM64

Built from main branch at commit 7ac8b49.

What's included

  • mlx-server — standalone OpenAI-compatible inference server
  • Standard MLX inference (mlx-lm)
  • DFlash speculative decoding acceleration

Usage

tar -xzf dflash-mlx-server-macos-arm64.tar.gz

# Standard MLX inference
./mlx-server --model /path/to/model --port 8080

# DFlash accelerated inference
./mlx-server --model /path/to/model --draft-model /path/to/draft --port 8080

For Atomic Chat integration

This binary is automatically downloaded during the Atomic Chat build process.

DFlash MLX Server macOS ARM64 (61b57ba)

Choose a tag to compare

DFlash MLX Server — macOS ARM64

Built from main branch at commit 61b57ba.

What's included

  • mlx-server — standalone OpenAI-compatible inference server
  • Standard MLX inference (mlx-lm)
  • DFlash speculative decoding acceleration

Usage

tar -xzf dflash-mlx-server-macos-arm64.tar.gz

# Standard MLX inference
./mlx-server --model /path/to/model --port 8080

# DFlash accelerated inference
./mlx-server --model /path/to/model --draft-model /path/to/draft --port 8080

For Atomic Chat integration

This binary is automatically downloaded during the Atomic Chat build process.

DFlash MLX Server macOS ARM64 (1936b7a)

Choose a tag to compare

DFlash MLX Server — macOS ARM64

Built from main branch at commit 1936b7a.

What's included

  • mlx-server — standalone OpenAI-compatible inference server
  • Standard MLX inference (mlx-lm)
  • DFlash speculative decoding acceleration

Usage

tar -xzf dflash-mlx-server-macos-arm64.tar.gz

# Standard MLX inference
./mlx-server --model /path/to/model --port 8080

# DFlash accelerated inference
./mlx-server --model /path/to/model --draft-model /path/to/draft --port 8080

For Atomic Chat integration

This binary is automatically downloaded during the Atomic Chat build process.

DFlash MLX Server macOS ARM64 (2c70d44)

Pre-release

Choose a tag to compare

DFlash MLX Server — macOS ARM64

Built from main branch at commit 2c70d44.

What's included

  • mlx-server — standalone OpenAI-compatible inference server
  • Standard MLX inference (mlx-lm)
  • DFlash speculative decoding acceleration

Usage

tar -xzf dflash-mlx-server-macos-arm64.tar.gz

# Standard MLX inference
./mlx-server --model /path/to/model --port 8080

# DFlash accelerated inference
./mlx-server --model /path/to/model --draft-model /path/to/draft --port 8080

For Atomic Chat integration

This binary is automatically downloaded during the Atomic Chat build process.

DFlash MLX Server macOS ARM64 (1ed2330)

Choose a tag to compare

DFlash MLX Server — macOS ARM64

Built from main branch at commit 1ed2330.

What's included

  • mlx-server — standalone OpenAI-compatible inference server
  • Standard MLX inference (mlx-lm)
  • DFlash speculative decoding acceleration

Usage

tar -xzf dflash-mlx-server-macos-arm64.tar.gz

# Standard MLX inference
./mlx-server --model /path/to/model --port 8080

# DFlash accelerated inference
./mlx-server --model /path/to/model --draft-model /path/to/draft --port 8080

For Atomic Chat integration

This binary is automatically downloaded during the Atomic Chat build process.

DFlash MLX Server macOS ARM64 (89d59f0)

Choose a tag to compare

DFlash MLX Server — macOS ARM64

Built from main branch at commit 89d59f0.

What's included

  • mlx-server — standalone OpenAI-compatible inference server
  • Standard MLX inference (mlx-lm)
  • DFlash speculative decoding acceleration

Usage

tar -xzf dflash-mlx-server-macos-arm64.tar.gz

# Standard MLX inference
./mlx-server --model /path/to/model --port 8080

# DFlash accelerated inference
./mlx-server --model /path/to/model --draft-model /path/to/draft --port 8080

For Atomic Chat integration

This binary is automatically downloaded during the Atomic Chat build process.

DFlash MLX Server macOS ARM64 (f486034)

Choose a tag to compare

DFlash MLX Server — macOS ARM64

Built from main branch at commit f486034.

What's included

  • mlx-server — standalone OpenAI-compatible inference server
  • Standard MLX inference (mlx-lm)
  • DFlash speculative decoding acceleration

Usage

tar -xzf dflash-mlx-server-macos-arm64.tar.gz

# Standard MLX inference
./mlx-server --model /path/to/model --port 8080

# DFlash accelerated inference
./mlx-server --model /path/to/model --draft-model /path/to/draft --port 8080

For Atomic Chat integration

This binary is automatically downloaded during the Atomic Chat build process.

DFlash MLX Server macOS ARM64 (b742c32)

Choose a tag to compare

DFlash MLX Server — macOS ARM64

Built from main branch at commit b742c32.

What's included

  • mlx-server — standalone OpenAI-compatible inference server
  • Standard MLX inference (mlx-lm)
  • DFlash speculative decoding acceleration

Usage

tar -xzf dflash-mlx-server-macos-arm64.tar.gz

# Standard MLX inference
./mlx-server --model /path/to/model --port 8080

# DFlash accelerated inference
./mlx-server --model /path/to/model --draft-model /path/to/draft --port 8080

For Atomic Chat integration

This binary is automatically downloaded during the Atomic Chat build process.