60. Alexandre Paris - Proofcheck - LLM fine-tuning and customization

60. Alexandre Paris - Proofcheck - LLM fine-tuning and customization

Author: Manuel Pasieka August 28, 2024 Duration: 53:19

## Summary

Today on the show I am talking to Proofreads CTO Alexandre Paris. Alex explains in great detail how they analyze digital books drafts to identify mistakes and instances within the document that dont follow guidelines and preferences of the user.


Alex is explaining how they fine-tune LLMs like, Mistrals 7B to achieve efficient resource usage and provide customizations and serve multiple uses cases with a single base model and multiple lora adapters.


We talk about the challenges and capabilities of fine-tuning, how to do it, when to apply it and when for example prompt engineering of an foundation model is the better choice.


I think this episode is very interesting for listeners that are using LLMs in a specific domain. It shows how fine-tuning a base model on selected high quality corpus can be used to build solutions outperform generic offerings by OpenAI or Google.


## AAIP Community

Join our discord server and ask guest directly or discuss related topics with the community.

https://discord.gg/5Pj446VKNU


## TOC

00:00:00 Beginning

00:02:46 Guest Introduction

00:06:12 Proofcheck Intro

00:11:43 Document Processing Pipeline

00:26:46 Customization Options

00:29:49 LLM fine-tuning

00:42:08 Prompt-engineering vs. fine-tuning


### References

https://www.proofcheck.io/

Alexandre Paris - https://www.linkedin.com/in/alexandre-paris-92446b22/


Hosted by Manuel Pasieka, the Austrian Artificial Intelligence Podcast offers a grounded, local perspective on a global phenomenon. Instead of abstract theorizing, each conversation focuses on the tangible impact and practical applications of AI within Austria's unique ecosystem. You'll hear from a diverse range of guests-researchers, entrepreneurs, policymakers, and creatives-who are actively shaping this landscape, discussing both the remarkable opportunities and the nuanced challenges specific to the region. The discussions delve into how these technologies are being integrated into Austrian industry, academia, and society, moving beyond hype to examine real-world implementation and ethical considerations. This podcast serves as an essential audio forum for anyone in Austria, or with an interest in the European tech scene, looking to understand how artificial intelligence is evolving right here. It’s about the people behind the algorithms and the local stories within a global revolution. For those engaged with the content, questions and suggestions are always welcome at the provided email address.
Author: Language: English Episodes: 73

Austrian Artificial Intelligence Podcast
Podcast Episodes
71 - NeoAlp - Humanoide Robotik - Zwischen Dystopie und Euphorie [not-audio_url] [/not-audio_url]

Duration: 1:41:48
Die letzten Jahre ware von Large Language Models dominiert, und für die meisten ist GenAI immer noch der Inbegriff von Fortschritt.Andere denken schon weiter und zielen auf physical AI ab, welche sich darauf spezialisier…
67. Mathias Neumayer and Dima Rubanov - Lora a child friendly AI [not-audio_url] [/not-audio_url]

Duration: 53:10
## SummaryLarge Language Models have many strengths and the frontier of what is possible and what they can be used for, is pushed back on the daily bases. One area in which current LLM's need to improve is how they commu…
66. Taylor Peer - Beat Shaper - A music producers AI Copilot [not-audio_url] [/not-audio_url]

Duration: 52:20
Today on the show I have the pleasure to talk to returning guest, Taylor Peer one of the co-founders of the startup, behind Beat Shaper.Taylor will explain how they are following an Bottom-up approach to create electroni…
65. Daniel Kondor - CSH - The long term impact of AI on society [not-audio_url] [/not-audio_url]

Duration: 1:04:50
Guest in this episode is the Computational Social Scientist Daniel Kondor, Postdoc at the Complexity Science Hub in Vienna.Daniel is talking about research methods that make it possible to study the impact of various fac…
64. Solo - Manuel Pasieka on the hottest LLM topics of 2024 [not-audio_url] [/not-audio_url]

Duration: 59:00
With the last episode in 2024, I dare to release an solo episode, summarizing my christmas research on the topics of - Small Language models - Agentic Systems - Advanced Reasoning / Test time compute paradigm I hope you…