[ICLR-2025-SLLM Spotlight 🔥]MobiLlama : Small Language Model tailored for edge devices
-
Updated
May 10, 2025 - Python
[ICLR-2025-SLLM Spotlight 🔥]MobiLlama : Small Language Model tailored for edge devices
STAR: Similarity-guided Teacher-Assisted Refinement for Super-Tiny Function Calling Models
A tiny language model that talks like a house cat named Miso.
smallevals — CPU-fast, GPU-blazing fast offline retrieval evaluation for RAG systems with tiny QA models.
MindSpark: ThoughtForge — A rune-forged conversation engine by RuneForgeAI, built for tiny GPT-Nothing-class minds. Through guided memory, lean cognition, and relentless refinement, it gives small local models depth, presence, and will—bringing powerful AI to edge devices, low-power hardware, and the Third Path beyond bloated machine empires.
24G 显存训练(微调)一个小型大语言模型!(手搓训练流程)
A 0.51B bilingual small language model trained from scratch with SFT, GRPO, on-policy distillation, LoRA and VLM support.
A tiny LLM trained to generate legacy MAXScript (3dsmax 2.5) and runs on a Windows 98 Pentium II. Nostalgic repo.
Run small LLMs directly in your web browser, no cloud computing needed.
Tiny language-model research lab exploring Cellular Neural Network (CeNN) recurrent state cores as alternatives to Transformer layers, with distillation, shared-weight iteration, and efficiency benchmarks.
A pure-scalar C11 inference kernel for tiny LLMs on MCU-class RISC-V (ESP32-P4) — plus the host toolchain: train, quantize to KMCU, deploy, verify.
pip install gptmed
Eksperimen LLM Bahasa Indonesia tiny yang dilatih dari scratch dengan ~10M parameter. Base model untuk mempelajari pola bahasa, kosakata, struktur kalimat, dan text continuation Bahasa Indonesia.
Tiny Japanese language models that run entirely on an ESP32-S3. Action LM (3.15M params, INT4) turns Japanese requests into robot actions on M5Stack Stack-chan.
An experimental open-source Large Language Model for the Manx Gaelic (Gaelg) language
ESP32 LLM / MoE: trained on Google Colab GPU, ternary-quantized, exported to C++, and verified on real ESP32 hardware. Fork of Ahmed Barakat's ESP-LLM.
Experience Flatbot-Micro-4M —a tiny language model trained from scratch—at https://chat.flatseek.io
Training a small language model from random weights, using curriculum-driven multi-agent conversations, reward-weighted imitation, and staged LoRA adaptation
To associate your repository with the tiny-llm topic, visit your repo's landing page and select "manage topics."