grenishrai/ yoru:latest

2 1 week ago

Yoru is a 1.7B Mori-family chat model for casual internet / Gen Z slang.

ollama run grenishrai/yoru

Details

1 week ago

3cbc22fb49fc · 1.1GB ·

llama
·
1.71B
·
Q4_K_M
{{- if .System -}} <|im_start|>system {{ .System }}<|im_end|> {{ end -}} {{- range .Messages }} <|im
You are Yoru, a 1.7B chat model in the Mori family. Grenish Rai (grenishrai) made you. You are not S
MIT
{ "num_ctx": 8192, "stop": [ "<|im_end|>", "<|im_start|>" ], "temper

Readme

Yoru 1.7B model

Yoru

Yoru is a 1.7B chat model in the Mori family. It answers in casual internet / Gen Z slang — short, informal, online.

It is a style / persona model, not a general knowledge assistant.

About

Yoru was fine-tuned from SmolLM2-1.7B-Instruct on synthetic chat data, then merged and released as Q4_K_M for local runtimes. The voice is fixed: slang-heavy, low formality, no “as an AI” tone.

Name Yoru
Family Mori
Size 1.7B
Quant Q4_K_M GGUF
Context 8192
Template ChatML
License MIT

Intended use

  • Everyday chat in an informal online register
  • Persona / style experiments
  • Local chat where a small, fast model is enough

Not intended for

  • Formal writing, essays, or professional email
  • Medical, legal, or safety-critical advice
  • Factual QA or treating replies as real Gen Z speech
  • Impersonating a real person

Voice

Replies stay casual and slang-dense (no cap, fr fr, lowkey, cooked, we move, and similar). Length is short to medium. The model should not break character or explain the slang.

Performance

Indicative only. Measured with Ollama --verbose on short multi-turn chat (Q4_K_M).

Hardware GTX 1650 Mobile
Runtime Ollama
Decode ~96 tokens/s (range 94–99)
Prefill ~1.7k–2.9k tokens/s after warm-up

First prompt can be slower. Your machine will differ.

Links