A ~303M-parameter, English-first causal language model — pretrained from scratch on the 4.5B-token LUNA_PreTrain corpus. The scaled-up sibling of LUNA-100M.
✓ TRAINED (pretraining complete) — final fp32 weights available at pretrained/final/lit_model.pth. Instruction tuning (RAG + MCP SFT, mirroring LUNA-100M) is planned on top of this checkpoint.