bupalinyu commited on
Commit
269b9a2
·
verified ·
1 Parent(s): cadc35c

Upload README.md

Browse files
Files changed (1) hide show
  1. README.md +44 -2
README.md CHANGED
@@ -7,6 +7,11 @@ tags:
7
  - prerouter
8
  - lora
9
  - ssd-offload
 
 
 
 
 
10
  base_model:
11
  - inclusionAI/Ling-3.0-tiny-base
12
  pipeline_tag: text-generation
@@ -22,6 +27,8 @@ pipeline_tag: text-generation
22
 
23
  **1 GiB active memory · 25 tok/s · 4-bit**
24
 
 
 
25
  [![GitHub](https://img.shields.io/badge/GitHub-Edge0--AI%2Fedge0-black?style=for-the-badge&logo=github)](https://github.com/Edge0-AI/edge0)
26
  [![Hugging Face](https://img.shields.io/badge/%F0%9F%A4%97%20Hugging%20Face-Edge0--35b--a3b--preview-yellow?style=for-the-badge)](https://huggingface.co/Edge0/Edge0-35b-a3b-preview)
27
  [![Hugging Face](https://img.shields.io/badge/%F0%9F%A4%97%20Hugging%20Face-Edge0--8b--a1b--preview-yellow?style=for-the-badge)](https://huggingface.co/Edge0/Edge0-8b-a1b-preview)
@@ -30,6 +37,11 @@ pipeline_tag: text-generation
30
  [![arXiv](https://img.shields.io/badge/arXiv-2609.18063-B31B1B?style=for-the-badge&logo=arxiv&logoColor=white)](https://arxiv.org/abs/2609.18063)
31
  [![License](https://img.shields.io/badge/License-Apache%202.0-blue?style=for-the-badge)](https://github.com/Edge0-AI/edge0/blob/main/LICENSE)
32
 
 
 
 
 
 
33
  </div>
34
 
35
 
@@ -53,8 +65,36 @@ via the [edge0](https://github.com/Edge0-AI/edge0) streaming inference framework
53
 
54
  </div>
55
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
56
  ## Highlights
57
 
 
 
 
 
 
58
  - **Runs in phone-class memory**: the full 4-bit checkpoint stays on
59
  storage and experts are streamed on demand, so only the active
60
  weights are in RAM — under **1 GiB**, with no sharding and no
@@ -143,8 +183,10 @@ Measured with `examples/bench.py` on a Mac mini M4 Pro, 24 GB:
143
  agentic tasks — tool use, multi-step planning, and long-horizon
144
  autonomy are currently weak. The full release will substantially
145
  strengthen agent capability.
146
- - The MLX backend currently targets Apple Silicon; other backends are
147
- on the edge0 roadmap.
 
 
148
 
149
  ## Quick start
150
 
 
7
  - prerouter
8
  - lora
9
  - ssd-offload
10
+ - multi-platform
11
+ - ios
12
+ - android
13
+ - windows
14
+ - macos
15
  base_model:
16
  - inclusionAI/Ling-3.0-tiny-base
17
  pipeline_tag: text-generation
 
27
 
28
  **1 GiB active memory · 25 tok/s · 4-bit**
29
 
30
+ **Native engines on iOS · macOS · Android · Windows.**
31
+
32
  [![GitHub](https://img.shields.io/badge/GitHub-Edge0--AI%2Fedge0-black?style=for-the-badge&logo=github)](https://github.com/Edge0-AI/edge0)
33
  [![Hugging Face](https://img.shields.io/badge/%F0%9F%A4%97%20Hugging%20Face-Edge0--35b--a3b--preview-yellow?style=for-the-badge)](https://huggingface.co/Edge0/Edge0-35b-a3b-preview)
34
  [![Hugging Face](https://img.shields.io/badge/%F0%9F%A4%97%20Hugging%20Face-Edge0--8b--a1b--preview-yellow?style=for-the-badge)](https://huggingface.co/Edge0/Edge0-8b-a1b-preview)
 
37
  [![arXiv](https://img.shields.io/badge/arXiv-2609.18063-B31B1B?style=for-the-badge&logo=arxiv&logoColor=white)](https://arxiv.org/abs/2609.18063)
38
  [![License](https://img.shields.io/badge/License-Apache%202.0-blue?style=for-the-badge)](https://github.com/Edge0-AI/edge0/blob/main/LICENSE)
39
 
40
+ [![iOS](https://img.shields.io/badge/iOS-000000?style=for-the-badge&logo=apple&logoColor=white)](https://github.com/Edge0-AI/edge0/tree/main/ios)
41
+ [![macOS](https://img.shields.io/badge/macOS-000000?style=for-the-badge&logo=apple&logoColor=white)](https://github.com/Edge0-AI/edge0/tree/main/macos)
42
+ [![Android](https://img.shields.io/badge/Android-3DDC84?style=for-the-badge&logo=android&logoColor=black)](https://github.com/Edge0-AI/edge0/tree/main/android)
43
+ [![Windows](https://img.shields.io/badge/Windows-0078D6?style=for-the-badge&logo=windows11&logoColor=white)](https://github.com/Edge0-AI/edge0/tree/main/windows)
44
+
45
  </div>
46
 
47
 
 
65
 
66
  </div>
67
 
68
+ ## Platforms — one model, four native engines
69
+
70
+ <div align="center">
71
+
72
+ **On 2026-09-30 we released the edge0 inference engines for
73
+ four platforms — so users get the best inference experience across
74
+ architectures and platforms. The source is open-sourced in the
75
+ [edge0 repo](https://github.com/Edge0-AI/edge0).**
76
+
77
+ | Platform | Native engine (open source) |
78
+ |---|---|
79
+ | 📱 **iOS** | [edge0/ios](https://github.com/Edge0-AI/edge0/tree/main/ios) |
80
+ | 🖥️ **macOS** | [edge0/macos](https://github.com/Edge0-AI/edge0/tree/main/macos) |
81
+ | 🤖 **Android** | [edge0/android](https://github.com/Edge0-AI/edge0/tree/main/android) |
82
+ | 🪟 **Windows** | [edge0/windows](https://github.com/Edge0-AI/edge0/tree/main/windows) |
83
+
84
+ </div>
85
+
86
+ This checkpoint is built for all of them: one model directory — the same
87
+ int4 base plus LoRA / prerouter adapters — runs unchanged on every
88
+ platform, so what you download here is what ships on a phone, a desktop
89
+ and a laptop alike.
90
+
91
  ## Highlights
92
 
93
+ - **Runs on iOS, macOS, Android and Windows**: the edge0 inference
94
+ engines are **open-source and native on all four platforms** — one
95
+ model, the best inference experience on every architecture and
96
+ platform (see [Platforms](#platforms--one-model-four-native-engines)
97
+ below).
98
  - **Runs in phone-class memory**: the full 4-bit checkpoint stays on
99
  storage and experts are streamed on demand, so only the active
100
  weights are in RAM — under **1 GiB**, with no sharding and no
 
183
  agentic tasks — tool use, multi-step planning, and long-horizon
184
  autonomy are currently weak. The full release will substantially
185
  strengthen agent capability.
186
+ - Platform coverage: the performance numbers above are measured with
187
+ the MLX backend on Apple Silicon; engine coverage and tuning on the
188
+ iOS / Android / Windows engines are still maturing (see
189
+ [Platforms](#platforms--one-model-four-native-engines)).
190
 
191
  ## Quick start
192