🌡️ It worked. Now I don't trust it.
Efficient small language models;
Architectures that enable small models to achieve capabilities comparable to much larger models;
Mixture-of-Experts and routing architectures;
Local LLM inference (vLLM, Vulkan-based inference);
Hierarchical and modular model memory;
Trainable memory modules that can be added, removed, or updated independently of the base model;
Continual learning and modular knowledge representation;