Papers
Communities
Events
Blog
Pricing
Search
Open menu
Home
Papers
2505.20622
Cited By
SeqPO-SiMT: Sequential Policy Optimization for Simultaneous Machine Translation
27 May 2025
Ting Xu
Zhichao Huang
Jiankai Sun
Shanbo Cheng
Wai Lam
OffRL
Re-assign community
ArXiv (abs)
PDF
HTML
Papers citing
"SeqPO-SiMT: Sequential Policy Optimization for Simultaneous Machine Translation"
1 / 1 papers shown
Title
DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
DeepSeek-AI
Daya Guo
Dejian Yang
Haowei Zhang
Junxiao Song
...
Shiyu Wang
S. Yu
Shunfeng Zhou
Shuting Pan
S.S. Li
ReLM
VLM
OffRL
AI4TS
LRM
390
2,024
0
22 Jan 2025
1