Wisp: small Llama-style base models trained from scratch on educational text and code.
DedeProGames PRO
DedeProGames
P(doom) <0.1%
AI & ML interests
Agentic & Coding finetune, Decoder-Onlys from Scratch, Image LoRAs, AI Researcher
Recent Activity
liked a model about 4 hours ago
convaiinnovations/laya repliedto Banaxi-Tech's post about 7 hours ago
Its SAICR time tomorrow.
Get ready!
https://huggingface.co/saicr repliedto their post about 8 hours ago
Im working on a 23M ASR model, trained on 100k hours of audioOrganizations
Kiyo
Kiyo: State-of-the-art models trained from scratch on vast amounts of web tokens, educational text, and code.
LowOnMind
LowOnMind: Small decoder-only models trained in small amount of tokens
NanoAndy
NTX
Chennus Series
Small and Efficient Chess Models
GPT-U
GPT-U: small Llama-style base models trained from scratch on web, educational text and code, with fully open scripts, logs and evals.
DynamicMind
DynamicMind: a family of lightweight models trained from scratch on diverse datasets.
Experimental Decoder-Onlys
Some experimental decoder-onlys made by me for research
NTX-2.1
medqwen
Experimental medical model based on Qwen2.5 Family
Wisp
Wisp: small Llama-style base models trained from scratch on educational text and code.
GPT-U
GPT-U: small Llama-style base models trained from scratch on web, educational text and code, with fully open scripts, logs and evals.
Kiyo
Kiyo: State-of-the-art models trained from scratch on vast amounts of web tokens, educational text, and code.
DynamicMind
DynamicMind: a family of lightweight models trained from scratch on diverse datasets.
LowOnMind
LowOnMind: Small decoder-only models trained in small amount of tokens
Experimental Decoder-Onlys
Some experimental decoder-onlys made by me for research
NanoAndy
NTX-2.1
NTX
medqwen
Experimental medical model based on Qwen2.5 Family
Chennus Series
Small and Efficient Chess Models