Inference Scaling for Long-Context RAG

Para podcasters

Spreaker Create

Nuestra plataforma

Noticias de productos

Registrate

Nuestra plataforma

Spreaker Create

Noticias de productos

Configuración

Tema claro

Tema oscuro

Inference Scaling for Long-Context RAG

20 de oct. de 2024 · 12m 17s

Inference Scaling for Long-Context RAG

Inference Scaling for Long-Context RAG

Descripción

🗓 Inference Scaling for Long-Context Retrieval Augmented Generation This research paper explores the effectiveness of inference scaling for retrieval augmented generation (RAG), a technique that enhances large language models (LLMs)...

mostra más

🗓 Inference Scaling for Long-Context Retrieval Augmented Generation

This research paper explores the effectiveness of inference scaling for retrieval augmented generation (RAG), a technique that enhances large language models (LLMs) by incorporating external knowledge. The authors introduce two strategies, demonstration-based RAG (DRAG) and iterative demonstration-based RAG (IterDRAG), for effectively scaling inference computation. They demonstrate that increasing inference computation, when optimally allocated, leads to nearly linear gains in RAG performance. Furthermore, they develop a computation allocation model to predict the optimal test-time compute allocation for various tasks and scenarios, showcasing its effectiveness in achieving performance gains and aligning with experimental results.

📎 Link to paper

mostra menos

Comentarios

Inicia sesión para dejar un comentario

Información

Autor	Shahriar Shariati
Organización	Shahriar Shariati
Página web	-
Etiquetas	#inference_scalling #llm #long_context

🇬🇧 English

🇮🇹 Italiano

🇪🇸 Espanõl

🇬🇧 English

🇮🇹 Italiano

🇪🇸 Espanõl

Copyright 2024 - Spreaker Inc. an iHeartMedia Company

Reproduciendo ahora Cola

Parece que no tienes ningún episodio activo

Echa un ojo al catálogo de Spreaker para descubrir nuevos contenidos.

Actual

Portada del podcast

Parece que no tienes ningún episodio en cola

Echa un ojo al catálogo de Spreaker para descubrir nuevos contenidos.

Siguiente

Portada del episodio

Portada del episodio

Portada del episodio

Portada del episodio

Cuánto silencio hay aquí...

¡Es hora de descubrir nuevos episodios!