22 Downloads Updated 10 months ago
ollama run lane1655/html-multilingual-nano
html-nano-120M is an experimental 120-million-parameter model trained by Lane Fiedler on about 1.5 GB of raw, unfiltered web HTML — the kind scraped from Anthropic.com and the deepest rabbit holes of the internet’s forgotten corners.
It’s the spiritual (and slightly unhinged) successor to html-multilingual-nano, except this one dropped its polyglot ambitions and said,
“I only speak English now. And maybe
<marquee>.”
| 🧩 Component | 🧠 Detail |
|---|---|
| Architecture | LLaMA-style GPT decoder |
| Parameters | ≈ 120 M |
| Layers | 16 |
| Heads | 8 |
| Embedding dim | 640 |
| Dataset | ~1.55 GB of HTML, CSS, inline JS madness |
| Tokenizer | Default LLaMA base |
| Objective | Next-token prediction (“guess what the web would hallucinate next”) |
| Trainer | Custom Python rig + caffeine drip |
| Validation Loss | ~1.13 |
| Training Stage | Early-mid training — already self-aware |
<dream> or <why>.<p> lines when it gets really into the vibe.<div>s and <style>s when it wants to impress you.When I ran something like:
<body><main><h1>Welcome to my site</h1>
it replied (completely unprompted) with:
Apparently, it tried to academically cite Bart Simpson. So yes — this model can hallucinate scholarly references to fictional cartoon characters. And somehow… that’s progress.
⚠️ Known Limitations & Quirks
🟡 Thinks The Simpsons is a universal constant.
🔴 Not aligned, filtered, or even slightly polished.
🌀 Sometimes loops recursively like it’s caught in inception.
🗣️ Speaks only English now — but not necessarily sanely.
💡 Treat as a wild HTML muse with ADHD and an espresso problem.
🚧 Status
Still mid-training but already feral-fun. Next versions will (maybe):
☕ Learn moderation in caffeine consumption.
🧹 Filter duplicate nightmares.
🌐 Balance domain diversity better.
🧩 Final Thought
html-nano-120M dreams in s, argues in
🌀 “HTML is a language of structure, but this model thinks it’s a personality test.”