Zzong's Notes
Search
검색
다크 모드
라이트 모드
탐색기
LLM
54건의 항목
2026년 9월 04일
Group Sequence Policy Optimization
LLM
reinforcement_learning
post_training
RLVR
policy_optimization
Qwen
2026년 9월 04일
BEQUE - Large Language Model based Long-tail Query Rewriting in Taobao Search
e-commerce
WWW
paper_review
query_rewriting
LLM
y2024
todo
trans
2026년 9월 04일
Rotary Positional Embedding
NLP
LLM
2026년 9월 04일
Prompt Engineering
LLM
NLP
2026년 9월 04일
RLHF
LLM
reinforcement_learning
2026년 9월 04일
Trainer(huggingface)
huggingface
LLM
2026년 9월 04일
TrainerCallback
huggingface
deep_learning
LLM
2026년 9월 04일
Batch Decoding
LLM
2026년 9월 04일
Prompt Compression
LLM
inference
prompt_compression
2026년 9월 04일
supervised fine-tuning
deep_learning
LLM
generative_ai
alignment
2026년 9월 04일
Chain of Hindsight Aligns Language Models with Feedback
language_model
LLM
nlp
paper_review
y2023
2026년 9월 04일
Decoupled Clip and Dynamic sAmpling Policy Optimization
LLM
reinforcement_learning
post_training
RLVR
policy_optimization
ByteDance
2026년 9월 04일
DPO
LLM
2026년 9월 04일
Group Relative Policy Optimization
reinforcement_learning
LLM
2026년 9월 04일
자연스러운 한국어를 위한 LLM Post-Training
LLM
post_training
alignment
instruction_tuning
korean
2026년 8월 29일
FastChat
LLM
2026년 8월 29일
Megatron-LM
LLM
distributed_training
2026년 8월 29일
polyglot
LLM
korean
NLP
2026년 8월 29일
DistributedDataParallel
deep_learning
LLM
2026년 8월 29일
FP8 Quantization
LLM
quantization
inference
serving
Qwen
2026년 8월 29일
Generalized Knowledge Distillation
LLM
distillation
2026년 8월 29일
GPT-2
LLM
2026년 8월 29일
GPT
NLP
LLM
2026년 8월 29일
LVLM 를 따라하는 정보 넣기
LLM
LVLM
2026년 8월 29일
Large Language Model
NLP
LLM
2026년 8월 29일
Llama
LLM
2026년 8월 29일
LoRA
LLM
2026년 8월 29일
Open challenges in LLM research
LLM
research
2026년 8월 29일
Proximal Policy Optimization
reinforcement_learning
LLM
2026년 8월 29일
T5
summarization
LLM
2026년 8월 29일
Train Large Model
LLM
deepspeed
2026년 8월 29일
deepspeed
deep_learning
LLM
2026년 8월 29일
KV Cache
LLM
inference
transformer
2026년 8월 29일
LLMLingua - Prompt Compression for LLM Inference
LLM
inference
prompt_compression
RAG
paper_review
y2023
y2024
2026년 8월 29일
Prefill
LLM
inference
serving
latency
2026년 8월 29일
Prompt Compression Trends - From Token Pruning to Context Engineering
LLM
inference
prompt_compression
context_compression
RAG
y2026
2026년 8월 29일
vllm
LLM
2026년 8월 29일
llm_as_classifier
LLM
2026년 8월 29일
On-Policy Distillation
LLM
distillation
alignment
generative_ai
2026년 8월 29일
vicuna
LLM
2026년 8월 29일
MT-bench
LMSYS
LLM
evaluation
2026년 8월 29일
DeepSpeed-MoE
MoE
deep_learning
LLM
2026년 8월 29일
LLMOps
MLOps
LLM
generative_model
server
inference
pipeline
2026년 8월 29일
Revisiting Reinforcement Learning for LLM Reasoning from A Cross-Domain Perspective
LLM
paper_review
2026년 8월 29일
Compact Language Models via Pruning andKnowledge Distillation
language_model
distillation
pruning
LLM
nlp
paper_review
sLLM
2026년 8월 29일
DialogLM
language_model
LLM
Microsoft
NLP
nlp
paper_review
summarization
y2022
2026년 8월 29일
LIMA
language_model
LLM
nlp
paper_review
2026년 8월 29일
Reinforcement Learning with Verifiable Rewards Implicitly Incentivizes Correct Reasoning in Base LLMs
LLM
RLVR
paper_review
reinforcement_learning
2026년 8월 29일
A Survey on AI Search with LargeLanguage Models
retrieval
IR
survey
LLM
2026년 8월 29일
LLM2Vec
embedding
dense_retrieval
LLM
2026년 8월 29일
GRAM - Generative Retrieval and Alignment Model
retrieval
e-commerce
generative_retrieval
LLM
paper_review
y2025
JD
2026년 8월 29일
LREF
retrieval
IR
paper_review
LLM
e-commerce
relevance
y2025
2026년 8월 29일
Towards More Relevant Product Search Ranking Via Large Language Models
retrieval
e-commerce
ranking
LLM
learning_to_rank
paper_review
2026년 8월 29일
Yelp - Search Query Understanding with LLMs
retrieval
e-commerce
query_understanding
LLM