Justin commited on
Commit
3acd574
·
verified ·
1 Parent(s): 979f2e7

Upload folder using huggingface_hub

Browse files
Files changed (5) hide show
  1. README.md +36 -0
  2. config.json +31 -0
  3. dit.safetensors +3 -0
  4. pos_emb.safetensors +3 -0
  5. vae.safetensors +3 -0
README.md ADDED
@@ -0,0 +1,36 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ ---
2
+ license: apache-2.0
3
+ base_model: ByteDance-Seed/SeedVR2-3B
4
+ base_model_relation: quantized
5
+ library_name: mlx-serve
6
+ tags:
7
+ - image-restoration
8
+ - super-resolution
9
+ - video-restoration
10
+ - mlx
11
+ ---
12
+
13
+ # SeedVR2-3B-MLX-Serve
14
+
15
+ [mlx-serve](https://github.com/justinluque/mlx-serve)-native repack of [ByteDance-Seed/SeedVR2-3B](https://huggingface.co/ByteDance-Seed/SeedVR2-3B) for one-step diffusion image/video restoration and upscaling.
16
+
17
+ Weights are a straight fp16 passthrough from the [Comfy-Org/SeedVR2](https://huggingface.co/Comfy-Org/SeedVR2) fp16 mirror (`diffusion_models/seedvr2_3b_fp16.safetensors`, `vae/ema_vae_fp16.safetensors`) plus the original repo's `pos_emb.pt` text-conditioning tensor — **no requantization**. Converted with mlx-serve's `tests/convert_seedvr2_weights.py`.
18
+
19
+ ## Files
20
+
21
+ - `config.json` — `model_type: seedvr2`
22
+ - `dit.safetensors` — the 635 NaDiT tensors (~6.8 GB fp16)
23
+ - `vae.safetensors` — encoder + decoder (~0.5 GB fp16)
24
+ - `pos_emb.safetensors` — the fixed text-conditioning tensor (~1 MB)
25
+
26
+ ## Usage
27
+
28
+ ```
29
+ mlx-serve --model-dir ~/.mlx-serve/models/ByteDance-Seed
30
+ ```
31
+
32
+ Serves `POST /v1/images/upscales` (images) and `POST /v1/video/upscales` (video frames).
33
+
34
+ ## License
35
+
36
+ Apache-2.0, inherited from the original ByteDance-Seed/SeedVR2-3B release.
config.json ADDED
@@ -0,0 +1,31 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ {
2
+ "model_type": "seedvr2",
3
+ "architectures": [
4
+ "NaDiT"
5
+ ],
6
+ "vid_dim": 2560,
7
+ "num_layers": 32,
8
+ "mm_layers": 10,
9
+ "heads": 20,
10
+ "head_dim": 128,
11
+ "vid_in_channels": 33,
12
+ "vid_out_channels": 16,
13
+ "patch_size": [
14
+ 1,
15
+ 2,
16
+ 2
17
+ ],
18
+ "rope_dim": 128,
19
+ "window": [
20
+ 4,
21
+ 3,
22
+ 3
23
+ ],
24
+ "vae": {
25
+ "latent_channels": 16,
26
+ "spatial_downsample_factor": 8,
27
+ "temporal_downsample_factor": 4,
28
+ "scaling_factor": 0.9152
29
+ },
30
+ "_converted_by": "tests/convert_seedvr2_weights.py"
31
+ }
dit.safetensors ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:2fd0e03a3dad24e07086750360727ca437de4ecd456f769856e960ae93e2b304
3
+ size 6783018808
pos_emb.safetensors ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:b171266bfa59ce32a1fcbf6d74eb874be84d12333e124631995a373d865f0283
3
+ size 1187920
vae.safetensors ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:20678548f420d98d26f11442d3528f8b8c94e57ee046ef93dbb7633da8612ca1
3
+ size 501324814