十六週Sixteen weeks
進度與進行方式Schedule and format
每週都有講解與實習討論時間。每週上課前,請儘量閱讀相關閱讀材料。Each week: lecture and discussion. Please read the related readings before each session.
W01 · 09/10
工具與環境:AI 帳號、Colab、編輯器Setup: AI accounts, Colab, editors
課程導論:兩種語言觀的對話Course intro: two ways of looking at language
這學期我們試著用不同角度來解讀 LLM:使用者、工程師與語言學家What perspectives do we read LLMs from: user, engineer, or linguist?
lab 1: 請在助教的協助下完成 AI 帳號 (Google Gemini) 學生免費方案申請,課堂上一起把計算環境確認。lab 1: Apply for the AI account (Google Gemini) student free tier during class; we confirm the computing environment together.
W02 · 09/17
NLP 與語言學的分合史NLP and linguistics: a history of divergence
模組化與背後的語言觀點Modularity and linguistic perspectives
從管線到端到端,NLP 丟掉了哪些語言學假設 ?Moving to end-to-end systems, which linguistic assumptions did NLP drop — and rightly so?
lab 2: TBAlab 2: TBA
W03 · 09/24
詞元化:BPE、WordPiece、SentencePieceTokenization: BPE, WordPiece, SentencePieceA1 作業A1 作業
語言單位與構式語法Linguistic units and Construction Grammar
詞元不是詞、也不是語素 — 那它到底是什麼?A token is not a word, a morpheme. So what is it?
lab 3: TBAlab 3: TBA
W04 · 10/01
向量表徵:從 word2vec 到脈絡化嵌入Embeddings: from word2vec to contextual representations
語意表徵是什麼意思 ?Semantic representation
分布假說能多大程度的表徵語意?剩下的部分是誰在分擔?How much meaning can distribution carry, and who does the rest?
lab 4: TBAlab 4: TBA
W05 · 10/08
Transformer、微調與對齊Transformers, fine-tuning and alignmentA1 繳交A1 dueA2 作業A2 作業
語料庫、句法與言談分析Corpora, syntax and discourse analysis
RLHF 訓練出來的語域是誰的語域?對齊算不算是一種語言規範化?評測工具是什麼?Whose register does RLHF produce? Is alignment a form of linguistic standardization? What are the evaluation tools?
lab 5: TBAlab 5: TBA
W06 · 10/15
語言工程:prompt、context、harness、loops 與 graph engineeringLanguage Re-engineering: Prompts, context, harness, loops and graph engineering
自然語言作為程式語言:vibe coding 與 skill writingNatural language as programming language: vibe coding and skill writing
大型語言模型正在重新工程化語言,使其從表徵世界的媒介,轉化為協調認知與行動的基礎層。LLMs are re-engineering language from a medium of representation into an orchestration layer for cognition and action.
lab 6: TBAlab 6: TBA
W07 · 10/22
推理:CoT、測試階段計算與推理模型Reasoning: chain-of-thought, test-time compute, reasoning models
語言與推理:把思考說出來為什麼有效Language and reasoning: why thinking out loud helps
把思考說出來為什麼會讓答案變準?那還算是思考嗎?Why does thinking out loud improve answers — and is it still thinking?
lab 7: TBAlab 7: TBA
W08 · 10/29
記憶:脈絡視窗、RAG 與長期記憶Memory: context windows, RAG, long-term memoryA2 繳交A2 dueA3 作業A3 作業
語言與記憶:從認知科學看人類與機器記憶Language and memory: from cognitive science to human and machine memory
檢索式記憶像不像人的記憶?人類回憶其實是重建而非調閱。Is retrieval like human memory? Human recall reconstructs rather than retrieves.
lab 8: TBAlab 8: TBA
W09 · 11/05
多模態(一):語音模型Multimodal I: speech models
語音、韻律與聽覺感知Speech, prosody and auditory perception
語音模型是先辨音再解義,還是根本沒有音位這一層?Do speech models recognize sounds then meanings, or is there no phonemic level?
lab 9: TBAlab 9: TBA
W10 · 11/12
多模態(二):視覺與跨模態對齊Multimodal II: vision and cross-modal alignment
多模態語言與指涉Multimodal language and reference
圖文對齊學到的是指涉關係,還是只是共現統計?Does image–text alignment learn reference, or only co-occurrence?
ICAIF 11/14–17;本週前後的截止日已預留緩衝。ICAIF Nov 14–17; deadlines around this week have buffer.
W11 · 11/19
具身 AI 與世界模型Embodied AI and world modelsA3 繳交A3 dueA4 作業A4 作業
具身認知與本體論Embodied cognition and ontologies
沒有身體的模型,能不能懂「重」「近」「痛」這種詞?Can a model without a body understand heavy, near, painful?
W12 · 11/26
小模型與本地 AI:蒸餾與量化Small models and local AI: distillation and quantization期末提案Proposal due
推理加速、效率與語言不平等Inference efficiency and linguistic inequality
當運算變便宜,低資源語言會被照顧到,還是被更快地拋下?As compute gets cheaper, are low-resource languages served or left behind faster?
教師日本東大參訪 11/25–28,本週助教上課。Instructor travelling Nov 25–28. Assistant teaching this week.
W13 · 12/03
多代理人(一):溝通與協作Multi-agent systems I: communication and collaboration
語用學與協定;語言、決策與行動Pragmatics and protocols; language, decision and action
代理人之間的協定需不需要合作原則?還是它們會發明自己的?Do agent protocols need Gricean cooperation, or will agents invent their own?
NeurIPS 12/6–12,本週課後不另安排 office hours。NeurIPS Dec 6–12; no extra office hours this week.
W14 · 12/10
多代理人(二):治理、安全與評估Multi-agent systems II: governance, safety and evaluationA4 繳交A4 due
人機社會語言學、多輪對話與代理人語言Human–AI sociolinguistics, multi-turn dialogue, emergent languages
跟模型講話久了,是誰被誰影響?After enough turns with a model, who accommodates to whom?
PACLIC 12/10–12,本週非同步或安排客座講者。PACLIC Dec 10–12. Asynchronous session or guest speaker.
W15 · 12/17
機制可解釋性(一):電路與特徵Mechanistic interpretability I: circuits and features
LLM 神經語言學The neurolinguistics of LLMs
在模型裡找到一條主謂一致的電路,算不算找到了語法?If we find an agreement circuit inside a model, have we found grammar?
W16 · 12/24
機制可解釋性(二):安全稽核與功能性情緒Mechanistic interpretability II: safety audits and functional emotions期末發表Presentations
語言、情緒與意識Language, emotion and consciousness
若模型內部有一個穩定對應「焦慮」的表徵,我們該怎麼描述它?If a stable internal representation tracks anxiety, how should we describe it?
期末發表;書面報告 Week 17 前上傳。Final presentations; written reports due Week 17.