Sebastian Raschka
|
dc2f8e95d4
|
Support different Qwen3 sizes in pkg (#714)
|
2025-06-28 08:00:23 -05:00 |
|
Sebastian Raschka
|
81be5fab0b
|
Add Qwen3 1.7, 4B, 8B, and 32B support to from-scratch nb (#709)
|
2025-06-25 08:53:09 -05:00 |
|
Sebastian Raschka
|
58b30e2f7b
|
Improve KV cache code for torch.compile (#705)
* Improve KV cache code for torch.compile
* cleanup
* cleanup
|
2025-06-23 18:08:49 -05:00 |
|
Sebastian Raschka
|
e9ffdbace4
|
CPU compile performance for Qwen3 models (#704)
* Ch06 classifier function asserts
* Qwen3 cpu compilation perf
|
2025-06-23 11:06:10 -05:00 |
|
Sebastian Raschka
|
0b15a00574
|
Qwen3 KV cache (#688)
|
2025-06-21 17:34:39 -05:00 |
|
Sebastian Raschka
|
9f6f514191
|
Fix formatting in Qwen3 nb (#680)
* Fix formatting in Qwen3 nb
* upd
|
2025-06-20 07:28:27 -05:00 |
|
Sebastian Raschka
|
3d4bce6d57
|
Qwen3 From Scratch (#678)
* Qwen3 From Scratch
* rev other file
* upd
* upd
* upd
* url fixes
|
2025-06-19 18:44:38 -05:00 |
|