oobabooga
|
6e8fb0e7b1
|
Update llama.cpp
|
2025-12-14 13:32:14 -08:00 |
|
oobabooga
|
9fe40ff90f
|
Update exllamav3 to 0.0.18
|
2025-12-10 05:37:33 -08:00 |
|
oobabooga
|
8e762e04b4
|
Merge remote-tracking branch 'refs/remotes/origin/dev' into dev
|
2025-12-09 05:27:43 -08:00 |
|
oobabooga
|
aa16266c38
|
Update llama.cpp
|
2025-12-09 03:19:23 -08:00 |
|
oobabooga
|
502f59d39b
|
Update diffusers to 0.36
|
2025-12-08 05:08:54 -08:00 |
|
oobabooga
|
e7c8b51fec
|
Revert "Use flash_attention_2 by default for Transformers models"
This reverts commit 85f2df92e9.
|
2025-12-07 18:48:41 -08:00 |
|
oobabooga
|
b758059e95
|
Revert "Clear the torch cache between sequential image generations"
This reverts commit 1ec9f708e5.
|
2025-12-07 12:23:19 -08:00 |
|
oobabooga
|
1ec9f708e5
|
Clear the torch cache between sequential image generations
|
2025-12-07 11:49:22 -08:00 |
|
oobabooga
|
3b8369a679
|
Update llama.cpp
|
2025-12-07 11:18:36 -08:00 |
|
oobabooga
|
058e78411d
|
docs: Small changes
|
2025-12-07 10:16:08 -08:00 |
|
oobabooga
|
17bd8d10f0
|
Update exllamav3 to 0.0.17
|
2025-12-07 09:37:18 -08:00 |
|
oobabooga
|
85f2df92e9
|
Use flash_attention_2 by default for Transformers models
|
2025-12-07 06:56:58 -08:00 |
|
oobabooga
|
1762312fb4
|
Use random instead of np.random for image seeds (makes it work on Windows)
|
2025-12-06 20:10:32 -08:00 |
|
oobabooga
|
160a25165a
|
docs: Small change
|
2025-12-06 08:41:12 -08:00 |
|
oobabooga
|
f93cc4b5c3
|
Add an API example to the image generation tutorial
|
2025-12-06 08:33:06 -08:00 |
|
oobabooga
|
c026dbaf64
|
Fix API requests always returning the same 'created' time
|
2025-12-06 08:23:21 -08:00 |
|
oobabooga
|
194e4c285f
|
Update llama.cpp
|
2025-12-06 08:14:48 -08:00 |
|
oobabooga
|
1c36559e2b
|
Add a News section to the README
|
2025-12-06 07:05:00 -08:00 |
|
oobabooga
|
02518a96a9
|
Lint
|
2025-12-06 06:55:06 -08:00 |
|
oobabooga
|
0100ad1bd7
|
Add user_data/image_outputs to the Gradio allowed paths
|
2025-12-06 06:39:30 -08:00 |
|
oobabooga
|
6411142111
|
docs: Small changes
|
2025-12-06 06:36:16 -08:00 |
|
oobabooga
|
455dc06db0
|
Serve the original PNG images in the UI instead of webp
|
2025-12-06 05:43:00 -08:00 |
|
oobabooga
|
1a9ed1fe98
|
Fix the height of the image output gallery
|
2025-12-06 05:21:26 -08:00 |
|
oobabooga
|
17b12567d8
|
docs: Small changes
|
2025-12-05 14:15:15 -08:00 |
|
oobabooga
|
e20b2d38ff
|
docs: Add VRAM measurements for Z-Image-Turbo
|
2025-12-05 14:12:08 -08:00 |
|
oobabooga
|
6ca99910ba
|
Image: Quantize the text encoder for lower VRAM
|
2025-12-05 13:08:46 -08:00 |
|
oobabooga
|
11937de517
|
Use flash attention for image generation by default
|
2025-12-05 12:13:24 -08:00 |
|
oobabooga
|
eba8a59466
|
docs: Improve the image generation tutorial
|
2025-12-05 12:10:41 -08:00 |
|
oobabooga
|
5848c7884d
|
Increase the height of the image output gallery
|
2025-12-05 10:24:51 -08:00 |
|
oobabooga
|
c11c14590a
|
Image: Better LLM variation default prompt
|
2025-12-05 08:08:11 -08:00 |
|
oobabooga
|
0dd468245c
|
Image: Add back the gallery cache (for performance)
|
2025-12-05 07:11:38 -08:00 |
|
oobabooga
|
b63d57158d
|
Image: Add TGW as a prefix to output images
|
2025-12-05 05:59:54 -08:00 |
|
oobabooga
|
afa29b9554
|
Image: Several fixes
|
2025-12-05 05:58:57 -08:00 |
|
oobabooga
|
8eac99599a
|
Image: Better LLM variation default prompt
|
2025-12-04 19:58:06 -08:00 |
|
oobabooga
|
b4f06a50b0
|
fix: Pass bos_token and eos_token from metadata to jinja2
Fixes loading Seed-Instruct-36B
|
2025-12-04 19:11:31 -08:00 |
|
oobabooga
|
15c6e43597
|
Image: Add a revised_prompt field to API results for OpenAI compatibility
|
2025-12-04 17:41:09 -08:00 |
|
oobabooga
|
56f2a9512f
|
Revert "Image: Add the LLM-generated prompt to the API result"
This reverts commit c7ad28a4cd.
|
2025-12-04 17:34:27 -08:00 |
|
oobabooga
|
3ef428efaa
|
Image: Remove llm_variations from the API
|
2025-12-04 17:34:17 -08:00 |
|
oobabooga
|
c7ad28a4cd
|
Image: Add the LLM-generated prompt to the API result
|
2025-12-04 17:22:08 -08:00 |
|
oobabooga
|
b451bac082
|
Image: Improve a log message
|
2025-12-04 16:33:46 -08:00 |
|
oobabooga
|
47a0fcd614
|
Image: PNG metadata improvements
|
2025-12-04 16:25:48 -08:00 |
|
oobabooga
|
ac31a7c008
|
Image: Organize the UI
|
2025-12-04 15:45:04 -08:00 |
|
oobabooga
|
a90739f498
|
Image: Better LLM variation default prompt
|
2025-12-04 10:50:40 -08:00 |
|
oobabooga
|
ffef3c7b1d
|
Image: Make the LLM Variations prompt configurable
|
2025-12-04 10:44:35 -08:00 |
|
oobabooga
|
5763947c37
|
Image: Simplify the API code, add the llm_variations option
|
2025-12-04 10:23:00 -08:00 |
|
oobabooga
|
2793153717
|
Image: Add LLM-generated prompt variations
|
2025-12-04 08:10:24 -08:00 |
|
oobabooga
|
7fb9f19bd8
|
Progress bar style improvements
|
2025-12-04 06:20:45 -08:00 |
|
oobabooga
|
a838223d18
|
Image: Add a progress bar during generation
|
2025-12-04 05:49:57 -08:00 |
|
oobabooga
|
14dbc3488e
|
Image: Clear the torch cache after generation, not before
|
2025-12-04 05:32:58 -08:00 |
|
oobabooga
|
235b94f097
|
Image: Add placeholder file for user_data/image_models
|
2025-12-03 18:43:30 -08:00 |
|
oobabooga
|
c357eed4c7
|
Image: Remove the flash_attention_3 option (no idea how to get it working)
|
2025-12-03 18:40:34 -08:00 |
|
oobabooga
|
c93d27add3
|
Update llama.cpp
|
2025-12-03 18:29:43 -08:00 |
|
oobabooga
|
fbca54957e
|
Image generation: Yield partial results for batch count > 1
|
2025-12-03 16:13:07 -08:00 |
|
oobabooga
|
49c60882bf
|
Image generation: Safer image uploading
|
2025-12-03 16:07:51 -08:00 |
|
oobabooga
|
59285d501d
|
Image generation: Small UI improvements
|
2025-12-03 16:03:31 -08:00 |
|
oobabooga
|
373baa5c9c
|
UI: Minor image gallery improvements
|
2025-12-03 14:45:02 -08:00 |
|
oobabooga
|
906dc54969
|
Load --image-model before --model
|
2025-12-03 12:15:38 -08:00 |
|
oobabooga
|
4468c49439
|
Add semaphore to image generation API endpoint
|
2025-12-03 12:02:47 -08:00 |
|
oobabooga
|
5ad174fad2
|
docs: Add an image generation API example
|
2025-12-03 11:58:54 -08:00 |
|
oobabooga
|
5433ef3333
|
Add an API endpoint for generating images
|
2025-12-03 11:50:56 -08:00 |
|
oobabooga
|
9448bf1caa
|
Image generation: add torchao quantization (supports torch.compile)
|
2025-12-02 14:22:51 -08:00 |
|
oobabooga
|
97281ff831
|
UI: Fix an index error in the new image gallery
|
2025-12-02 11:20:52 -08:00 |
|
oobabooga
|
9d07d3a229
|
Make portable builds functional again after b3666e140d
|
2025-12-02 10:06:57 -08:00 |
|
oobabooga
|
6291e72129
|
Remove quanto for now (requires messy compilation)
|
2025-12-02 09:57:18 -08:00 |
|
oobabooga
|
a83821e941
|
Revert "UI: Optimize typing in all textareas"
This reverts commit e24ba92ef2.
|
2025-12-01 10:34:23 -08:00 |
|
oobabooga
|
24fd963c38
|
Merge remote-tracking branch 'refs/remotes/origin/dev' into dev
|
2025-12-01 08:06:08 -08:00 |
|
oobabooga
|
e24ba92ef2
|
UI: Optimize typing in all textareas
|
2025-12-01 08:05:21 -08:00 |
|
oobabooga
|
78b315344a
|
Update exllamav3
|
2025-11-28 06:45:05 -08:00 |
|
oobabooga
|
3cad0cd4c1
|
Update llama.cpp
|
2025-11-28 03:52:37 -08:00 |
|
oobabooga
|
8f0048663d
|
More modular HTML generator
|
2025-11-21 07:09:16 -08:00 |
|
oobabooga
|
b0baf7518b
|
Remove macos x86-64 portable builds (macos-13 runner deprecated by GitHub)
|
2025-11-19 06:07:15 -08:00 |
|
oobabooga
|
0d4eff284c
|
Add a --cpu-moe model for llama.cpp
|
2025-11-19 05:23:43 -08:00 |
|
oobabooga
|
d6f39e1fef
|
Add ROCm portable builds
|
2025-11-18 16:32:20 -08:00 |
|
oobabooga
|
327a234d23
|
Add ROCm requirements.txt files
|
2025-11-18 16:24:56 -08:00 |
|
oobabooga
|
4e4abd0841
|
Merge remote-tracking branch 'refs/remotes/origin/dev' into dev
|
2025-11-18 14:07:05 -08:00 |
|
oobabooga
|
c45f35ccc2
|
Remove the macos 13 wheels (deprecated by GitHub)
|
2025-11-18 14:06:42 -08:00 |
|
oobabooga
|
d85b95bb15
|
Update llama.cpp
|
2025-11-18 14:06:04 -08:00 |
|
oobabooga
|
a26e28bdea
|
Update exllamav3 to 0.0.15
|
2025-11-18 11:24:16 -08:00 |
|
oobabooga
|
6a3bf1de92
|
Update exllamav3 to 0.0.14
|
2025-11-09 19:43:53 -08:00 |
|
oobabooga
|
e7534a90d8
|
Update llama.cpp
|
2025-11-05 18:46:01 -08:00 |
|
oobabooga
|
6be1bfcc87
|
Remove the CUDA 11.7 portable builds
|
2025-11-05 05:45:10 -08:00 |
|
oobabooga
|
92d9cd36a6
|
Update llama.cpp
|
2025-11-05 05:43:34 -08:00 |
|
oobabooga
|
67f9288891
|
Pin huggingface-hub to 0.36.0 (solves #7284 and #7289)
|
2025-11-02 14:01:00 -08:00 |
|
oobabooga
|
16f77b74c4
|
Merge remote-tracking branch 'refs/remotes/origin/dev' into dev
|
2025-11-01 19:58:53 -07:00 |
|
oobabooga
|
cd645f80f8
|
Update exllamav3 to 0.0.12
|
2025-11-01 19:58:18 -07:00 |
|
oobabooga
|
338ae36f73
|
Add weights_only=True to torch.load in Training_PRO
|
2025-10-28 12:43:16 -07:00 |
|
oobabooga
|
f4c9e67155
|
Update llama.cpp
|
2025-10-23 08:19:32 -07:00 |
|
oobabooga
|
24fd2b4dec
|
Update exllamav3 to 0.0.11
|
2025-10-21 07:26:38 -07:00 |
|
oobabooga
|
be81f050a7
|
Merge remote-tracking branch 'refs/remotes/origin/dev' into dev
|
2025-10-20 19:43:36 -07:00 |
|
oobabooga
|
9476123ee6
|
Update llama.cpp
|
2025-10-20 19:43:26 -07:00 |
|
oobabooga
|
a156ebbf76
|
Lint
|
2025-10-15 13:15:01 -07:00 |
|
oobabooga
|
c871d9cdbd
|
Revert "Same as 7f06aec3a1 but for exllamav3_hf"
This reverts commit deb37b821b.
|
2025-10-15 13:05:41 -07:00 |
|
oobabooga
|
163d863443
|
Update llama.cpp
|
2025-10-15 11:23:10 -07:00 |
|
oobabooga
|
c93d567f97
|
Update exllamav3 to 0.0.10
|
2025-10-15 06:41:09 -07:00 |
|
oobabooga
|
b5a6904c4a
|
Make --trust-remote-code immutable from the UI/API
|
2025-10-14 20:47:01 -07:00 |
|
oobabooga
|
efaf2aef3d
|
Update exllamav3 to 0.0.9
|
2025-10-13 15:32:25 -07:00 |
|
oobabooga
|
047855c591
|
Update llama.cpp
|
2025-10-13 15:32:03 -07:00 |
|
oobabooga
|
611399e089
|
Update README
|
2025-10-11 17:22:48 -07:00 |
|
oobabooga
|
968c79db06
|
Minor README fix (closes #7251)
|
2025-10-11 17:20:49 -07:00 |
|
oobabooga
|
655c3e86e3
|
Fix "continue" missing an initial space in chat-instruct/chat modes
|
2025-10-11 17:00:25 -07:00 |
|