Angaben zum Datum
Datum aus der Quelle.
Erstmals gesehen am .
Rapid-MLX 0.15.7: Embeddings und zuverlässigerer Agent-Setup
Rapid-MLX 0.15.7 liefert über /v1/embeddings native Text- und Code-Embeddings mit EmbeddingGemma 2 (768 Dimensionen, 8192-Token-Limit), verbessert die Einrichtung von Continue, Cline und Qwen Code und bringt Fixes für Prompt-Wiederverwendung sowie MTP-Scheduling.
What's new in v0.15.7
Rapid-MLX 0.15.7 adds native text and code embeddings, improves local agent setup, and strengthens Desktop and inference reliability.
Highlights
- EmbeddingGemma 2 text/code embeddings: serve normalized 768-dimensional vectors through
/v1/embeddings, with standard 4-bit and BF16 aliases, optional smaller dimensions, and an explicit 8192-token limit. This release supports text/code inputs; images, audio and video are unsupported. (#4279) - More reliable agent setup: Continue and Cline configuration is written where their clients read it; keyed Continue connections preserve the configured server credential. Qwen Code previews redact retained secrets while keeping the saved configuration intact. (#4257, #4274, #4275)
- Warmer coding-agent turns: prompt reuse avoids repeated full-prompt tokenization; MTP scheduling and reproducibility receive further fixes. Improvements depend on the model, hardware and workload. (#4214, #4221, #4238) …