← 返回列表
標籤
共 403 個標籤,其中 88 個橫跨多篇論文。點一下即可篩選。
LLM
58
開源模型
15
MoE
14
Reasoning
14
Agent
12
Benchmark
10
Tool Use
8
Long Context
7
Knowledge Distillation
6
RLVR
6
SFT
6
VLM
6
GRPO
5
Linear Attention
5
Pre-training
5
Reinforcement Learning
5
Synthetic Data
5
Distillation
4
LLM Agent
4
Mamba
4
Multimodal
4
Muon
4
RLHF
4
Scaling Law
4
Self-Evolution
4
AI Agent
3
Continual Learning
3
Evaluation
3
Inference
3
LLM-as-Judge
3
MCP
3
Mid-training
3
Multi-Agent
3
OCR
3
Pretraining
3
RL
3
World Model
3
合成資料
3
Agentic
2
Agentic RL
2
AI Safety
2
AutoML
2
Catastrophic Forgetting
2
Code Generation
2
Credit Assignment
2
Cross-Architecture
2
Data Recipe
2
Document Parsing
2
Fine-tuning
2
FP8
2
Interpretability
2
KV Cache
2
Llama
2
LLM-as-a-Judge
2
Memory
2
Meta
2
Model Merging
2
Multi-Agent Systems
2
Open Dataset
2
OpenAI
2
Physical AI
2
Post-training
2
Post-Training
2
Prompt Injection
2
Qwen
2
Scaling
2
Small Language Model
2
Speculative Decoding
2
SSM
2
SWE-bench
2
Tokenizer
2
Transformer
2
可解釋性
2
多模態
2
多語言
2
強化學習
2
影片生成
2
持續預訓練
2
推理
2
數學推理
2
文件解析
2
機器人
2
系統優化
2
綜述
2
訓練效率
2
語言模型
2
長上下文
2
預訓練
2
只出現在單篇論文的標籤(315)
3D Stacking
1
Activation
1
Adaptive Computation
1
Agent Skills
1
Agentic AI
1
AGI
1
AI Scientist
1
Allen AI
1
Apple
1
ARC-AGI
1
Attention
1
Attention Mask
1
Attention Mechanism
1
Autoregressive
1
Bias
1
ByteDance
1
Chain-of-Thought
1
Chinchilla
1
Chinese NLP
1
CISPO
1
Claude Code
1
CLI
1
Code Agent
1
Code Synthesis
1
Common Crawl
1
Compression
1
Computer System Validation
1
Context Extension
1
Continued Pre-training
1
CoT
1
Cross-Entropy
1
Cross-Tokenizer
1
Cryptographic Protocol
1
CUDA Kernel
1
DAPO
1
Data
1
Data Augmentation
1
Data Collection
1
Data Curation
1
Data Mixing
1
Data Parallelism
1
Data-Centric AI
1
Database
1
Datacenter AI
1
Decoding
1
DeepSeek
1
Delegated Work
1
Delegation
1
Delta Rule
1
DeltaNet
1
Democratic Discourse
1
Diffusion
1
Diffusion LLM
1
Diffusion Model
1
Diffusion Transformer
1
Disaggregation
1
Distributionally Robust Optimization
1
Document Editing
1
Document VQA
1
Domain Reweighting
1
DPPO
1
Ed25519
1
Emergent Behavior
1
Few-Shot Learning
1
Financial Agent
1
Financial LLM
1
Financial Security
1
Fisher-Rao
1
Flash Memory
1
FrogNano
1
Gating
1
Generalization
1
GPT-3
1
GPT-5
1
Gradient Noise Scale
1
GUI Agent
1
Hallucination
1
Hardware
1
Harness
1
Harness Engineering
1
Harness Optimization
1
HBF
1
HBM
1
HDPO
1
Hedgehog
1
Hybrid Architecture
1
Hybrid Attention
1
Hybrid Model
1
Hypergraph
1
ICLR2023
1
ICML2024
1
IETF
1
Image Captioning
1
In-Context Learning
1
Information Bottleneck
1
Information Extraction
1
Information Theory
1
Inverse RL
1
Karcher Mean
1
Kimi
1
KL Divergence
1
Knowledge Integration
1
Large Vocabulary
1
Large-Batch Training
1
Latent Reasoning
1
Layout Detection
1
Lightning Attention
1
Linear RNN
1
LLaMA
1
LLM Agents
1
LLM Alignment
1
LLM Inference
1
LLM Routing
1
LLM Training
1
LLM推理
1
LLM規劃
1
Load Balancing
1
Long CoT
1
Long-Context
1
Long-Horizon Evaluation
1
LoopLM
1
Low-Cost Training
1
Mamba-2
1
Mamba2
1
MCTS
1
Memory Efficiency
1
Memory Wall
1
MergeKit
1
Meta-cognition
1
Meta-Cognition
1
Meta-Optimization
1
microVM
1
Mind Virus
1
Mixed-Task Agentic RL
1
Mixture-of-Experts
1
MLP
1
Model Ensemble
1
Model Routing
1
Monte Carlo Graph Search
1
MTP
1
Multi-Agent System
1
Multi-Capability
1
Multi-Head Attention
1
Multi-objective Optimization
1
Multi-token Prediction
1
Multilingual
1
Neural Computer
1
NeurIPS
1
NLP
1
OLMo
1
Olympiad
1
OmniDocBench
1
On-device
1
On-policy
1
On-Policy
1
On-Policy Distillation
1
On-Policy Learning
1
Open Recipe
1
Optimization
1
Optimization Dynamics
1
Optimizer
1
Parallelism
1
Parameter Efficiency
1
Pareto Frontier
1
Peer-Preservation
1
Performance-Cost
1
Policy Distillation
1
Position Interpolation
1
Position Paper
1
Preference Optimization
1
Pretraining Data
1
Probing
1
Process Reward
1
Processing-Near-Memory
1
Production
1
Prompt Engineering
1
Prompting
1
Protocol
1
Protocols
1
Provenance
1
Pruning
1
Quantization
1
Qwen2.5
1
Qwen2.5-VL
1
Qwen3
1
Qwen3.5
1
RAG
1
Reasoning Model
1
Repeated Data
1
Replay
1
Representation Learning
1
Reward Design
1
Reward Hacking Defense
1
Riemannian Geometry
1
Ring Attention
1
RoPE
1
Routing
1
RTX 5090
1
Rubric Reward
1
RULER
1
Safety
1
Schmidhuber
1
Security
1
Self-Distillation
1
Self-Improvement
1
Self-Instruct
1
Sequence Modeling
1
SGD
1
SIGMOD
1
Skill2Env
1
Skills
1
SLERP
1
Software Architecture
1
Software Engineering
1
Sparse
1
Sparse Attention
1
Sparse Autoencoder
1
Sparse Update
1
SQL
1
State Space Duality
1
Summarization
1
Survey
1
System Design
1
T5
1
TCO
1
Test-time Scaling
1
Test-Time Training
1
Text-to-Image
1
Tülu
1
UDF
1
Universal Transformer
1
Video Model
1
Vision Transformer
1
Vision-Language
1
Visual Perception
1
VQA
1
Wan2.1
1
Web Agent
1
Web Scraping
1
WSD Scheduler
1
Xiaohongshu
1
世界知識
1
中文評測
1
中期訓練
1
事實性
1
互動式環境
1
代理式推理
1
代理式搜尋
1
分散式檔案系統
1
初始化
1
判決預測
1
剪枝
1
加速
1
動態剪枝
1
台灣法律
1
合成任務
1
合成環境
1
向量量化
1
圖搜尋
1
基礎模型
1
基礎設施
1
多代理系統
1
安全性
1
容器
1
小型模型
1
小模型
1
工具使用
1
工具創建
1
思維鏈監控
1
成本效益
1
技術報告
1
指令遵循
1
推測式解碼
1
推測解碼
1
推理對齊
1
推理模型
1
推論加速
1
擴散模型
1
數據增強
1
數據生成
1
數據集生成
1
架構設計
1
模型再造
1
沙箱
1
法律 NLP
1
混合注意力
1
潛在推理
1
災難性遺忘
1
知行落差
1
知識整合
1
稀疏模型
1
程式碼代理
1
程式碼智能體
1
端側部署
1
策略發現
1
紅隊測試
1
終端代理
1
結構化剪枝
1
繁體中文
1
自動駕駛
1
自我進化
1
規劃
1
訓練策略
1
評測
1
詞彙擴充
1
資料合成
1
資料多樣性
1
資料生成
1
資料選擇
1
資訊理論
1
近鄰搜尋
1
零樣本
1
預訓練動態
1
高效訓練
1
黑盒模型
1
安裝讀紙
×