Quartz 4
Search
Search
Dark mode
Light mode
Explorer
Home
❯
02 Sections
❯
CS224N
Folder: 02---Sections/CS224N
81 items under this folder.
Sep 24, 2026
CS224N 2026 - Lecture 12 - Reasoning Part 1
cs224n
lecture
Sep 24, 2026
CS224N 2026 - Lecture 13 - Reasoning Part 2
cs224n
lecture
Sep 24, 2026
CS224N 2026 - Lecture 14 - Tokenization and Multilinguality
cs224n
lecture
Sep 24, 2026
CS224N 2026 - Lecture 16 - AIs Impact on Humanity
cs224n
lecture
Sep 24, 2026
CS224N 2026 - Lecture 19 - The Art of Artificial Reasoning for Small Language Models
cs224n
lecture
Sep 24, 2026
SLP 2026 - Chapter 02 - Words and Tokens
cs224n
textbook
Sep 24, 2026
SLP 2026 - Chapter 09 - Post-training - Instruction Tuning Alignment and Test-Time
cs224n
textbook
Sep 24, 2026
SLP 2026 - Chapter 10 - Masked Language Models
cs224n
textbook
Sep 24, 2026
2025 - We Cant Understand AI Using Our Existing Vocabulary
cs224n
paper
Sep 24, 2026
CS224N - Notes - Backpropagation Old
cs224n
course-note
Sep 24, 2026
CS224N 2017 - Review of Differential Calculus Theory
cs224n
course-note
Sep 24, 2026
CS224N 2019 - Notes - Computing Neural Network Gradients
cs224n
course-note
Sep 24, 2026
CS224N 2019 - Notes 02 - Word Vectors II - GloVe Evaluation and Training
cs224n
course-note
Sep 24, 2026
CS224N 2019 - Notes 03 - Neural Networks and Backpropagation
cs224n
course-note
Sep 24, 2026
CS224N 2019 - Notes 05 - Language Models RNN GRU and LSTM
cs224n
course-note
Sep 24, 2026
CS224N 2023 - Notes 01 - Introduction and Word2Vec - Draft
cs224n
course-note
Sep 24, 2026
CS224N 2023 - Notes 10 - Self-Attention and Transformers - Draft
cs224n
course-note
Sep 24, 2026
CS224N 2026 - Lecture 02 - Word Vectors
cs224n
lecture
Sep 24, 2026
CS224N 2026 - Lecture 03 - Neural Network Foundations
cs224n
lecture
Sep 24, 2026
CS224N 2026 - Lecture 04 - Language Models and Recurrent Neural Networks
cs224n
lecture
Sep 24, 2026
CS224N 2026 - Lecture 05 - Attention and Transformers
cs224n
lecture
Sep 24, 2026
CS224N 2026 - Lecture 06 - Final Projects and Practical Tips
cs224n
lecture
Sep 24, 2026
CS224N 2026 - Lecture 07 - Pretraining
cs224n
lecture
Sep 24, 2026
CS224N 2026 - Lecture 08 - Post-training
cs224n
lecture
Sep 24, 2026
CS224N 2026 - Lecture 09 - Efficient Adaptation
cs224n
lecture
Sep 24, 2026
CS224N 2026 - Lecture 10 - RAG and Language Agents
cs224n
lecture
Sep 24, 2026
CS224N 2026 - Lecture 11 - Evaluation
cs224n
lecture
Sep 24, 2026
2011 - Natural Language Processing Almost from Scratch
cs224n
paper
Sep 24, 2026
2013 - Distributed Representations of Words and Phrases and their Compositionality
cs224n
paper
Sep 24, 2026
2013 - Efficient Estimation of Word Representations in Vector Space
cs224n
paper
Sep 24, 2026
2013 - On the Difficulty of Training Recurrent Neural Networks
cs224n
paper
Sep 24, 2026
2014 - GloVe - Global Vectors for Word Representation
cs224n
paper
Sep 24, 2026
2015 - Improving Distributional Similarity with Lessons Learned from Word Embeddings
cs224n
paper
Sep 24, 2026
2016 - A Latent Variable Model Approach to PMI-based Word Embeddings
cs224n
paper
Sep 24, 2026
2016 - Layer Normalization
cs224n
paper
Sep 24, 2026
2016 - Neural Machine Translation of Rare Words with Subword Units
cs224n
paper
Sep 24, 2026
2017 - Attention Is All You Need
cs224n
paper
Sep 24, 2026
2017 - Derivatives Backpropagation and Vectorization - Justin Johnson
cs224n
paper
Sep 24, 2026
2018 - BERT - Pre-training of Deep Bidirectional Transformers for Language Understanding
cs224n
paper
Sep 24, 2026
2018 - Image Transformer
cs224n
paper
Sep 24, 2026
2018 - Music Transformer - Generating Music with Long-Term Structure
cs224n
paper
Sep 24, 2026
2018 - On the Dimensionality of Word Embedding
cs224n
paper
Sep 24, 2026
2019 - Parameter-Efficient Transfer Learning for NLP
cs224n
paper
Sep 24, 2026
2019 - The Lottery Ticket Hypothesis - Finding Sparse Trainable Neural Networks
cs224n
paper
Sep 24, 2026
2020 - Contextual Word Representations - A Contextual Introduction
cs224n
paper
Sep 24, 2026
2020 - Language Models are Few-Shot Learners
cs224n
paper
Sep 24, 2026
2020 - Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks
cs224n
paper
Sep 24, 2026
2020 - Unsupervised Cross-lingual Representation Learning at Scale
cs224n
paper
Sep 24, 2026
2021 - LoRA - Low-Rank Adaptation of Large Language Models
cs224n
paper
Sep 24, 2026
2021 - Measuring Massive Multitask Language Understanding
cs224n
paper
Sep 24, 2026
2021 - RoFormer - Enhanced Transformer with Rotary Position Embedding
cs224n
paper
Sep 24, 2026
2022 - Chain-of-Thought Prompting Elicits Reasoning in Large Language Models
cs224n
paper
Sep 24, 2026
2022 - Fast Inference from Transformers via Speculative Decoding
cs224n
paper
Sep 24, 2026
2022 - Retrieval-Augmented Multimodal Language Modeling
cs224n
paper
Sep 24, 2026
2022 - Scaling Instruction-Finetuned Language Models
cs224n
paper
Sep 24, 2026
2023 - AlpacaFarm - A Simulation Framework for Methods that Learn from Human Feedback
cs224n
paper
Sep 24, 2026
2023 - Direct Preference Optimization - Your Language Model is Secretly a Reward Model
cs224n
paper
Sep 24, 2026
2023 - Do All Languages Cost the Same - Tokenization in the Era of Commercial Language Models
cs224n
paper
Sep 24, 2026
2023 - Holistic Evaluation of Language Models
cs224n
paper
Sep 24, 2026
2023 - How Far Can Camels Go - Exploring Instruction Tuning on Open Resources
cs224n
paper
Sep 24, 2026
2023 - Lets Verify Step by Step
cs224n
paper
Sep 24, 2026
2023 - ReAct - Synergizing Reasoning and Acting in Language Models
cs224n
paper
Sep 24, 2026
2023 - Scaling Autoregressive Multi-Modal Models - Pretraining and Instruction Tuning
cs224n
paper
Sep 24, 2026
2023 - Scaling Laws for Generative Mixed-Modal Language Models
cs224n
paper
Sep 24, 2026
2023 - Self-Consistency Improves Chain of Thought Reasoning in Language Models
cs224n
paper
Sep 24, 2026
2023 - Toolformer - Language Models Can Teach Themselves to Use Tools
cs224n
paper
Sep 24, 2026
2024 - Chameleon - Mixed-Modal Early-Fusion Foundation Models
cs224n
paper
Sep 24, 2026
2024 - LMFusion - Adapting Pretrained Language Models for Multimodal Generation
cs224n
paper
Sep 24, 2026
2024 - Scaling LLM Test-Time Compute Optimally Can Be More Effective Than Scaling Model Parameters
cs224n
paper
Sep 24, 2026
2024 - The Llama 3 Herd of Models
cs224n
paper
Sep 24, 2026
2024 - Transfusion - Predict the Next Token and Diffuse Images with One Multi-Modal Model
cs224n
paper
Sep 24, 2026
2024 - Visual Sketchpad - Sketching as a Visual Chain of Thought for Multimodal Language Models
cs224n
paper
Sep 24, 2026
2025 - Agentic Interpretability - Because We Have LLMs We Can and Should Pursue It
cs224n
paper
Sep 24, 2026
2025 - Bridging the Human-AI Knowledge Gap Through Concept Discovery and Transfer in AlphaZero
cs224n
paper
Sep 24, 2026
2025 - DAPO - An Open-Source LLM Reinforcement Learning System at Scale
cs224n
paper
Sep 24, 2026
2025 - DeepSeek-R1 - Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
cs224n
paper
Sep 24, 2026
2025 - Mixture-of-Transformers - A Sparse and Scalable Architecture for Multi-Modal Foundation Models
cs224n
paper
Sep 24, 2026
2025 - Multimodal RewardBench - Holistic Evaluation of Reward Models for Vision Language Models
cs224n
paper
Sep 24, 2026
2025 - Neologism Learning for Controllability and Self-Verbalization
cs224n
paper
Sep 24, 2026
2025 - OneFlow - Concurrent Mixed-Modal and Interleaved Generation with Edit Flows
cs224n
paper
Sep 24, 2026
2025 - Reconstruction Alignment Improves Unified Multimodal Models
cs224n
paper