LLMs-Journey
The LLMs-Journey repository covers topics including:
- Large Language Models (LLMs)
- Agentic Systems and Workflows
- Fine-Tuning (e.g., LoRA, QLoRa)
- Retrieval-Augmented Generation (RAG)
- Vision-Language Models (VLMs)
- Complementary Resources and Research
Blog Posts
Fine-Tuning:
Agents:
VLMs:
On-Device AI:
Misc:
What to read?
Books:
Agents:
Promptning:
Fine-Tuning:
RAG related:
Evaluation:
Datasets:
Models:
Vision:
Chain-of-Thought
Visualisation-of-Thought:
Test-Time Scaling
Test-Time Compute
AlphaGeometry:
Apple:
Misc:
- Scaling Pre-training to One Hundred Billion Data for Vision Language Models
- Competitive Programming with Large Reasoning Models
- MoBA: Mixture of Block Attention for Long-Context LLMs
- Step-Video-T2V Technical Report: The Practice, Challenges, and Future of Video Foundation Model, repo
- Native Sparse Attention: Hardware-Aligned and Natively Trainable Sparse Attention
- Deepseek Papers
- InfiniteHiP: Extending Language Model Context Up to 3 Million Tokens on a Single GPU
- Leveraging the true depth of LLMs
- NOLIMA: Long-Context Evaluation Beyond Literal Matching
- Memory Layers at Scale
- Towards an AI co-scientist
- On the consistent reasoning paradox of intelligence and optimal trust in AI: The power of ‘I don’t know’
- LIMO: Less is More for Reasoning
- How new data permeates LLM knowledge and how to dilute it
- Attention Is All You Need
- mHC: Manifold-Constrained Hyper-Connections
- GDPO: Group reward-Decoupled Normalization Policy Optimization for Multi-reward RL Optimization, GitHub
- Attention Residuals, GitHub
- TurboQuant, TurboQuant: Online Vector Quantization with Near-optimal Distortion Rate, QJL: 1-Bit Quantized JL Transform for KV Cache Quantization with Zero Overhead, PolarQuant: Quantizing KV Caches with Polar Transformation
- DSpark: Confidence-Scheduled Speculative Decoding with Semi-Autoregressive Generation, DeepSpec
- Sakana Fugu Technical Report, Blog
- TRINITY: AN EVOLVED LLM COORDINATOR, GitHub
- Reinforcement learning towards broadly and persistently beneficial models, Blog
- Qwen-Robot Suite: A Foundation Model Suite for Physical World Intelligence
Resources
Apple/MLX:
LangChain:
Vector Databases:
HuggingFace:
Leonie Notebooks:
GitHub Repos:
Blogs/Posts:
Frameworks: