🧭 Decide
🗂️ Sort
📊 Rank
💬 LLM chat
Input
Result
Pick a preset or write your own, then press Decide.
Sort items into categories
Each item is classified independently.
Result
Results grouped by category appear here.
Rank items by a criterion
Items are ordered by P(yes).
Result
Ranked list appears here.
LLM chat local · CPU · streaming
⚡ Stream demos (long output, watch tokens arrive)
⏳ Await demos (short answer, returned in one piece)
Uses the API key saved in this browser (🔑 top right). Stream shows tokens live (SSE). Await waits for the full reply. CPU only: Gemma 2B ~18 tok/s, 8K-token context; one request at a time (others queue). Thinking toggles only apply to reasoning models — Gemma 2B answers directly, so they are greyed out. Prefer Open WebUI pointed at
https://theetay.com/v1.Reply
statusidlefirst token–tokens/s–tokens–total–
💭 Thinking
No thinking yet.
Pick a preset or type a message, choose Stream or Await, then Send.