Sobes.tech

Machine Learning / AI

Как работал self-correction в Text-to-SQL агенте?

188

How was the inference of LLM (DeepSeek R1 33B) organized? How many GPUs, and how was parallelization handled?

187

Tell about the architecture of the RAG-chat (iChat) in TechVille: pipeline, components, technologies.

182

Tell me about the Text-to-SQL agent: architecture, pipeline, technologies.

179

What metrics were used to evaluate the RAG system and how was the decision made to deploy it to production?

177

How was the quality of the Text-to-SQL agent evaluated?

174

Tell us about the report summarization project at Alrosa: task, approach, fine-tuning.

169

Why did you choose Qdrant instead of staying with Elasticsearch? What are the advantages?

167

How were load peaks handled in the RAG system? What happened under high load?

163

What ranking metrics were used to evaluate the retriever?

153