Blacad Docs
拓展课 II LLM
正在初始化搜索引擎
    Blacad/blog
    • 首页
    • CS
    Blacad/blog
    • 首页
    • CS
      • CS61A
        • L1 Welcome
        • L2 Functions
        • L3 Control
        • L4 Higher Order Function
        • L5 Environments
        • L6 Sounds
        • L7 Function Abstract
        • L8 Function Example
        • L9 Recursion
        • L10 Tree Recursion
        • L11 Sequences
        • L12 Containers
        • L13 Data Abstraction
        • L14 Trees
        • L15 Mutability
        • L16 Iterators
        • L17 Generators
        • L18 Objects
        • L19 Class Attributes
        • L20 Inheritance
        • L21 Representations
        • L22 Composition
        • L23 Efficiency
        • L24 Decomposition
        • L25 Data Example
        • L28 Scheme
        • L29 Scheme Lists
        • L30 Calculator
        • L31 Interpreters
        • L32 Tail Calls
        • L33 Programs as Data
        • L34 Macros
        • L35 SQL
        • L36 Tables
        • L37 Aggregation
        • L38 Database
        • L39 思考解决问题
        • Hog
        • Cats
        • Ants
        • Scheme
        • Note
      • CS70
        • Note10 Counting
        • Note11 Countability
        • Note12 Computability
        • Note13 Discrete Probability
        • Note14 Conditional Probability
        • Note15 RV
        • Note16 RV
        • Note17 Concentration Inequalities and the Laws of Large Numbers
        • Note18 Hashing and Load Balance
        • Note19 Geometric and Poisson Distributions
        • Note20 Continuous Probability Distribution
        • Note21 Markov Chains
        • 不完备补充
        • 限与界
      • CS188
        • L1 Introduction
        • L2 Search
        • L3 Informed Search
        • L4 Local Search
        • L5 Games
        • L6 Uncertainty
        • L7 Logic
        • L8 Propositional Logic
        • L9 First Order Logic
        • L10 Probability
        • L11 Bayes Nets
        • L12 Bayes Nets Exact Inference
        • L13 Bayes Nets Approximate Inference
        • L14 Markov Models
        • L15 HMMS
        • L16 Dynamic Bayes Nets and Particle Filters
        • L17 Rational Decisions
        • L18 Markov Decision Process
        • L19 MDP II
        • L20 ML I
        • L21 ML II
        • L22 ML III
        • L23 ML IV
        • L24 RL I
        • L25 RL II
        • L26 RL III
        • 拓展课 I AI Existential Safety(RL)
        • 拓展课 II LLM

    拓展课 II LLM

    • LLM
      • 组成
        • Feature engineer
          • Text tokenization
          • Word embeddings
        • Deep neural networks
          • Autoregressive models
          • self-attention mechanisms
          • Transformer architecture
          • Multi-class classification
      • 技术
        • 训练
          • Supervised Learning
            • Self-supervised learning
            • Instruction tuning
          • Reinforcement learning
            • RLHF
        • Policy Search --- Proximal Policy Optimization
          • policy gradient methds
        • Beam Search