Thinking LLMs

Para podcasters

Spreaker Create

Nuestra plataforma

Noticias de productos

Registrate

Nuestra plataforma

Spreaker Create

Noticias de productos

Configuración

Tema claro

Tema oscuro

Thinking LLMs

18 de oct. de 2024 · 19m 33s

Thinking LLMs

Thinking LLMs

Descripción

🤔 Thinking LLMs: General Instruction Following with Thought Generation This research paper explores the concept of "Thinking LLMs," or large language models that can generate internal thoughts before responding to...

mostra más

🤔 Thinking LLMs: General Instruction Following with Thought Generation

This research paper explores the concept of "Thinking LLMs," or large language models that can generate internal thoughts before responding to user prompts. The authors propose a training method called Thought Preference Optimization (TPO) which uses an iterative process to encourage LLMs to develop thinking abilities. TPO leverages an existing judge model that evaluates responses, implicitly guiding the model to improve its thoughts based on the quality of the resulting responses. The study demonstrates that Thinking LLMs can outperform standard LLMs on various general instruction-following tasks, including those not typically associated with reasoning, such as marketing and health. The research highlights the potential for Thinking LLMs to expand the capabilities of these models beyond traditional reasoning and problem-solving domains.

📎 Link to paper

mostra menos

Comentarios

Inicia sesión para dejar un comentario

Información

Autor	Shahriar Shariati
Organización	Shahriar Shariati
Página web	-
Etiquetas	#planning #reasoning #tpo

🇬🇧 English

🇮🇹 Italiano

🇪🇸 Espanõl

🇬🇧 English

🇮🇹 Italiano

🇪🇸 Espanõl

Copyright 2024 - Spreaker Inc. an iHeartMedia Company

Reproduciendo ahora Cola

Parece que no tienes ningún episodio activo

Echa un ojo al catálogo de Spreaker para descubrir nuevos contenidos.

Actual

Portada del podcast

Parece que no tienes ningún episodio en cola

Echa un ojo al catálogo de Spreaker para descubrir nuevos contenidos.

Siguiente

Portada del episodio

Portada del episodio

Portada del episodio

Portada del episodio

Cuánto silencio hay aquí...

¡Es hora de descubrir nuevos episodios!