PARA-ZK
Search
Search
Dark mode
Light mode
Explorer
resource/ai
35 items with this tag.
Jul 13, 2026
Sources (example stubs)
resource/ai
Jul 13, 2026
LLaMA - Open and Efficient Foundation Language Models
resource/ai
Jul 13, 2026
Language Models are Few-Shot Learners
resource/ai
Jul 13, 2026
Language Models are Unsupervised Multitask Learners
resource/ai
Jul 13, 2026
Learning Transferable Visual Models From Natural Language Supervision
resource/ai
Jul 13, 2026
Llama 2 - Open Foundation and Fine-Tuned Chat Models
resource/ai
Jul 13, 2026
LoRA - Low-Rank Adaptation of Large Language Models
resource/ai
Jul 13, 2026
Mamba - Linear-Time Sequence Modeling with Selective State Spaces
resource/ai
Jul 13, 2026
Mixtral of Experts
resource/ai
Jul 13, 2026
QLoRA - Efficient Finetuning of Quantized LLMs
resource/ai
Jul 13, 2026
Qwen3 Technical Report
resource/ai
Jul 13, 2026
Qwen3-Coder-Next Technical Report
resource/ai
Jul 13, 2026
Retentive Network - A Successor to Transformer for Large Language Models
resource/ai
Jul 13, 2026
Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks
resource/ai
Jul 13, 2026
Scaling Laws for Neural Language Models
resource/ai
Jul 13, 2026
The Llama 3 Herd of Models
resource/ai
Jul 13, 2026
Training Compute-Optimal Large Language Models
resource/ai
Jul 13, 2026
Training language models to follow instructions with human feedback
resource/ai
Jul 13, 2026
Transformers are SSMs - Generalized Models and Efficient Algorithms Through Structured State Space Duality
resource/ai
Jul 13, 2026
mHC - Manifold-Constrained Hyper-Connections
resource/ai
Jul 13, 2026
An Image is Worth 16x16 Words - Transformers for Image Recognition at Scale
resource/ai
Jul 13, 2026
Attention Is All You Need
resource/ai
Jul 13, 2026
BERT - Pre-training of Deep Bidirectional Transformers for Language Understanding
resource/ai
Jul 13, 2026
Chain-of-Thought Prompting Elicits Reasoning in Large Language Models
resource/ai
Jul 13, 2026
DeepSeek-R1 - Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
resource/ai
Jul 13, 2026
DeepSeek-V3 Technical Report
resource/ai
Jul 13, 2026
DeepSeek-V3.2 - Pushing the Frontier of Open Large Language Models
resource/ai
Jul 13, 2026
Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer
resource/ai
Jul 13, 2026
FlashAttention - Fast and Memory-Efficient Exact Attention with IO-Awareness
resource/ai
Jul 13, 2026
GLM-4.5 - Agentic, Reasoning, and Coding (ARC) Foundation Models
resource/ai
Jul 13, 2026
GPT-4 Technical Report
resource/ai
Jul 13, 2026
Improving Language Understanding by Generative Pre-Training
resource/ai
Jul 13, 2026
Improving language models by retrieving from trillions of tokens
resource/ai
Jul 13, 2026
Kimi K2 - Open Agentic Intelligence
resource/ai
Jul 13, 2026
Kimi k1.5 - Scaling Reinforcement Learning with LLMs
resource/ai