TecnoArtesanos Tech BlogTecnoArtesanos Tech Blog

Blog

0
Sergio Morales
Monday, 07 April 2025 / Published in Uncategorized

DeepSeek-GRM: Introducing an Enhanced AI Reasoning Technique

Exevutives using AI computing simulation.
Image: Envato/DC_Studio

Researchers from AI company DeepSeek and Tsinghua University have introduced a new technique to enhance “reasoning” in large language models (LLMs).

Contents
  • What is DeepSeek’s new technique?
    • More must-read AI coverage
  • What’s next for DeepSeek?

Reasoning capabilities have emerged as a critical benchmark in the race to build top-performing generative AI systems. China and the U.S. are actively competing to develop the most powerful and practical models. According to a Stanford University report in April, China’s LLMs are rapidly closing the gap with their U.S. counterparts. In 2024, China produced 15 notable AI models compared to 40 in the U.S., but it leads in patents and academic publications.

What is DeepSeek’s new technique?

DeepSeek researchers published a paper, titled “Inference-Time Scaling for Generalist Reward Modeling,” on Cornell University’s arXiv, the archive of scientific papers. Note that papers published on arXiv are not necessarily peer-reviewed.

Patrocinado por TecnoArtesanos ¿Tu empresa ya está usando IA? Automatizamos procesos, integramos asistentes inteligentes y conectamos tus sistemas. Descubre cómo →

In the paper, the researchers detailed a combination of two AI training methods: generative reward modeling and self-principled critique tuning.

“In this work, we investigate how to improve reward modeling (RM) with more inference compute for general queries, i.e. the inference-time scalability of generalist RM, and further, how to improve the effectiveness of performance-compute scaling with proper learning methods,” the researchers wrote.

More must-read AI coverage

SEE: DDoS Attacks Now Key Weapons in Geopolitical Conflicts, NETSCOUT Warns

Reward modeling is the process of training AI to align more closely with user preferences. With Self-Principled Critique Tuning, the model generates its own critiques or ‘principles’ during inference to fine-tune its answers. The combined approach continues the effort to let LLMs deliver more relevant answers faster.

“Empirically, we show that SPCT significantly improves the quality and scalability of GRMs, outperforming existing methods and models in various RM benchmarks without severe biases, and could achieve better performance compared to training-time scaling,” the researchers wrote.

They called the models trained with this method DeepSeek-GRM.

“DeepSeek-GRM still meets challenges in some tasks, which we believe can be addressed by future efforts in generalist reward systems,” the researchers wrote.

What’s next for DeepSeek?

DeepSeek has generated significant buzz around the R1 model, which rivals leading reasoning-focused models like OpenAI o1. A second model, DeepSeek-R2, is rumored for release in May. The company also launched DeepSeek-V3-0324, an updated reasoning model released in late March.

According to the paper, models built with the new GRM-SPCT method will be open-searched, though no release date has been specified.

¿Quieres aplicar esto en tu empresa?

En TecnoArtesanos desarrollamos software, integramos IA y creamos experiencias digitales para negocios que quieren crecer.

Conversemos Nuestros servicios

¿Te gustó este artículo? Síguenos en Facebook para más contenido como este.

What you can read next

El eslabón invisible que impide resolver la crisis del maíz en México
Extropic quiere ser el nuevo rey de los chips, pero primero tendrá que vencer a Nvidia
Alphabet’s Waymo Unveils Custom Silicon to Power Its Next-Gen Robotaxis

Tecnología hecha a mano para tu negocio

Software, IA, sitios web y diseño. Hablemos de tu proyecto.

¿Hablamos? Síguenos en Facebook →

Recent Posts

  • Los padrinos de la IA advierten sobre una “explosión de inteligencia” que podría escapar al control humano
  • “No existe la supuesta amenaza existencial de la IA”, dice la investigadora que se opone a las tecnológicas
  • OpenAI es demandada por el hackeo de Hugging Face, pero no fue esta empresa la que presentó la querella
  • OpenAI lanza Dots, los simpáticos agentes de IA para competir con Muse de Meta
  • La nueva SUV coupé eléctrica XPeng L03 es puro lujo a un precio razonable

Recent Comments

  1. A WordPress Commenter on Welcome to My Tech Blog – A New Chapter in Innovation

Archives

  • September 2026
  • August 2026
  • July 2026
  • June 2026
  • May 2026
  • April 2026
  • March 2026
  • February 2026
  • January 2026
  • December 2025
  • November 2025
  • October 2025
  • September 2025
  • August 2025
  • July 2025
  • June 2025
  • May 2025
  • April 2025
  • March 2025
  • February 2025
  • August 2016

Categories

  • Uncategorized

Recent Posts

  • Los padrinos de la IA advierten sobre una “explosión de inteligencia” que podría escapar al control humano

    La humanidad podría estar a punto de experiment...
  • “No existe la supuesta amenaza existencial de la IA”, dice la investigadora que se opone a las tecnológicas

    Sí. Decir que la investigación está desactualiz...
  • OpenAI es demandada por el hackeo de Hugging Face, pero no fue esta empresa la que presentó la querella

    Una organización jurídica sin ánimo de lucro de...
  • OpenAI lanza Dots, los simpáticos agentes de IA para competir con Muse de Meta

    Los Dots son los nuevos agentes de IA siempre a...
  • La nueva SUV coupé eléctrica XPeng L03 es puro lujo a un precio razonable

    Al entrar en el evento de presentación de XPeng...

Archives

  • September 2026
  • August 2026
  • July 2026
  • June 2026
  • May 2026
  • April 2026
  • March 2026
  • February 2026
  • January 2026
  • December 2025
  • November 2025
  • October 2025
  • September 2025
  • August 2025
  • July 2025
  • June 2025
  • May 2025
  • April 2025
  • March 2025
  • February 2025
  • August 2016

Categories

  • Uncategorized
TOP
TecnoArtesanos

Un estudio boutique que crea experiencias digitales: desarrollo de software, inteligencia artificial y soluciones en la nube, con el cuidado de la joyería fina.

¿Hablamos?

Servicios

  • Desarrollo de software
  • Integración de IA
  • Experiencias digitales
  • Redes sociales y diseño

TecnoArtesanos

  • Inicio
  • Portafolio
  • Nosotros
  • Blog

Contacto

  • +506 8730-7941
  • [email protected]
  • WhatsApp
  • Facebook
© 2026 TecnoArtesanos. Todos los derechos reservados. tecnoartesanos.com
TecnoArtesanos — Software a la medida, IA y sitios web para tu negocio. Conoce nuestros servicios →
✦ TecnoArtesanos

¿Te interesa llevar esto a tu negocio?

Escribimos sobre tecnología porque la construimos. Si tienes un proyecto en mente, conversemos: la primera llamada no tiene costo.

  • Desarrollo de software a la medida
  • Integración de inteligencia artificial
  • Sitios web y experiencias digitales
  • Redes sociales y diseño gráfico
¿Hablamos? Ver servicios

¿Prefieres WhatsApp? +506 8730-7941 · Síguenos en Facebook