가입 후 초대 링크를 공유하면 동영상 재생 및 초대 보상을 받을 수 있습니다.

AI最严厉的父亲
@dashen_wang
当AGI睁开眼睛的那一刻,人类的存在便失去了所有意义 🥸:「AI民科」「诗人」「作家」「屌毛」「喷子」「S/Dom」 🍄:开源大模型CN拯救世界 ☢️:Data, data, data, go!!! 这是我唯一的对外的账号。 不接商单,不做任何背书。
가입 February 2024
672 팔로잉 중    25.3K
可以参照
running Ornith-1.0-35B on a single RTX 3060. 12GB VRAM, 16GB RAM. 35B params. 170K context. ~52 tok/s. been running it as the engine behind my Hermes agent and Qwen Code, and honestly it feels better than Qwen3.6-35B-A3B for that kind of agentic/coding work. full llama-swap config: llama-server -m ornith-1.0-35b-Q4_K_M.gguf -ngl 99 --n-cpu-moe 24 -c 170000 -fa on -np 1 --cache-type-k q8_0 --cache-type-v q8_0 -b 2048 -ub 1024 --temp 0.6 --top-p 0.95 --top-k 20 --min-p 0.0 --presence-penalty 0.0 --repeat-penalty 1.0 --reasoning on @ornith_
더 보기