CS
cs.AI
    All
  • CS
  • Economics
  • EES
  • Math
  • Physics
  • Biology
  • Finance
  • Statistics
  • All
  • cs.AI
  • cs.AR
  • cs.CC
  • cs.CE
  • cs.CG
  • cs.CL
  • cs.CR
  • cs.CV
  • cs.CY
  • cs.DB
  • cs.DC
  • cs.DL
  • cs.DM
  • cs.DS
  • cs.ET
  • cs.FL
  • cs.GL
  • cs.GR
  • cs.GT
  • cs.HC
  • cs.IR
  • cs.IT
  • cs.LG
  • cs.LO
  • cs.MA
  • cs.MM
  • cs.MS
  • cs.NA
  • cs.NE
  • cs.NI
  • cs.OH
  • cs.OS
  • cs.PF
  • cs.PL
  • cs.RO
  • cs.SC
  • cs.SD
  • cs.SE
  • cs.SI
  • cs.SY
  • All
  • econ.EM
  • econ.GN
  • econ.TH
  • All
  • eess.AS
  • eess.IV
  • eess.SP
  • eess.SY
  • All
  • math.AC
  • math.AG
  • math.AP
  • math.AT
  • math.CA
  • math.CO
  • math.CT
  • math.CV
  • math.DG
  • math.DS
  • math.FA
  • math.GM
  • math.GN
  • math.GR
  • math.GT
  • math.HO
  • math.IT
  • math.KT
  • math.LO
  • math.MG
  • math.MP
  • math.NA
  • math.NT
  • math.OA
  • math.OC
  • math.PR
  • math.QA
  • math.RA
  • math.RT
  • math.SG
  • math.SP
  • math.ST
  • All
  • astro-ph.CO
  • astro-ph.EP
  • astro-ph.GA
  • astro-ph.HE
  • astro-ph.IM
  • astro-ph.SR
  • cond-mat.dis-nn
  • cond-mat.mes-hall
  • cond-mat.mtrl-sci
  • cond-mat.other
  • cond-mat.quant-gas
  • cond-mat.soft
  • cond-mat.stat-mech
  • cond-mat.str-el
  • cond-mat.supr-con
  • gr-qc
  • hep-ex
  • hep-lat
  • hep-ph
  • hep-th
  • math-ph
  • nlin.AO
  • nlin.CD
  • nlin.CG
  • nlin.PS
  • nlin.SI
  • nucl-ex
  • nucl-th
  • physics.acc-ph
  • physics.ao-ph
  • physics.app-ph
  • physics.atm-clus
  • physics.atom-ph
  • physics.bio-ph
  • physics.chem-ph
  • physics.class-ph
  • physics.comp-ph
  • physics.data-an
  • physics.ed-ph
  • physics.flu-dyn
  • physics.gen-ph
  • physics.geo-ph
  • physics.hist-ph
  • physics.ins-det
  • physics.med-ph
  • physics.optics
  • physics.plasm-ph
  • physics.pop-ph
  • physics.soc-ph
  • physics.space-ph
  • quant-ph
  • All
  • q-bio.BM
  • q-bio.CB
  • q-bio.GN
  • q-bio.MN
  • q-bio.NC
  • q-bio.OT
  • q-bio.PE
  • q-bio.QM
  • q-bio.SC
  • q-bio.TO
  • All
  • q-fin.CP
  • q-fin.EC
  • q-fin.GN
  • q-fin.MF
  • q-fin.PM
  • q-fin.PR
  • q-fin.RM
  • q-fin.ST
  • q-fin.TR
  • All
  • stat.AP
  • stat.CO
  • stat.ME
  • stat.ML
  • stat.OT
  • stat.TH
Paper: Oct 13,2024
cs.AI
ID:2410.09671
OpenR: An Open Source Framework for Advanced Reasoning with Large Language Models
In this technical report, we introduce OpenR, an open-source framework designed to integrate key components for enhancing the reasoning capabilities of large language models (LLMs). OpenR unifies data acquisition, reinforcement learning training (both online and offline), and non-autoregressive decoding into a cohesive software platform. Our goal is to establish an open-source platform and community to accelerate the development of LLM reasoning. Inspired by the success of OpenAI's o1 model, which demonstrated improved reasoning abilities through step-by-step reasoning and reinforcement learning, OpenR integrates test-time compute, reinforcement learning, and process supervision to improve reasoning in LLMs. Our work is the first to provide an open-source framework that explores the core techniques of OpenAI's o1 model with reinforcement learning, achieving advanced reasoning capabilities beyond traditional autoregressive methods. We demonstrate the efficacy of OpenR by evaluating it on the MATH dataset, utilising publicly available data and search methods. Our initial experiments confirm substantial gains, with relative improvements in reasoning and performance driven by test-time computation and reinforcement learning through process reward models. The OpenR framework, including code, models, and datasets, is accessible at https://openreasoner.github.io. 🔗 View origin paper >>
Am I the publisher of this paper? I want to claim it
Claim
Paper Author: Jun Wang,Meng Fang,Ziyu Wan,Muning Wen,Jiachen Zhu,Anjie Liu,Ziqin Gong,Yan Song,Lei Chen,Lionel M. Ni,Linyi Yang,Ying Wen,Weinan Zhang
Leave an answer
Claim
Claim content
Report
Report content
Welcome!
or
*
Forgot password?
Don' have an account? Sign up