Ir al contenido principalSaltar al contenido

[7] [ arXiv:2608.28447 ](https://arxiv.org/abs/2608.28447 "Abstract") [[pdf](https://arxiv.org/pdf/2608.28447 "Download PDF"), [html](https://arxiv.org/html/2608.28447v1 "View HTML"), [other](https://

Abstract

[7] [ arXiv:2608.28447 ](https://arxiv.org/abs/2608.28447 "Abstract") [[pdf](https://arxiv.org/pdf/2608.28447 "Download PDF"), [html](https://arxiv.org/html/2608.28447v1 "View HTML"), [other](https://arxiv.org/format/2608.28447 "Other formats")] Title: Learning to Use Tools: Reinforcement Learning for Tool-Integrated Mathematical Reasoning [Minghui Xu](https://arxiv.org/search/cs?searchtype=author&query=Xu,+M), [Zi Wang](https://arxiv.org/search/cs?searchtype=author&query=Wang,+Z) Subject

Transparencia: Este análisis ha sido generado con asistencia de inteligencia artificial bajo supervisión editorial de SAPIENSDATAAI.

Cookies esenciales

Necesarias para el funcionamiento del sitio. No se pueden desactivar.

Cookies analíticas

Nos permiten medir el tráfico y mejorar el sitio (Google Analytics).

Más info: Política de Cookies