42. Rahim Entezari - TU-Graz & CSH - Improving generalization in parameter and data space

Author: Manuel Pasieka June 1, 2023 Duration: 1:05:27

Austrian Artificial Intelligence Podcast

Technology

# Summary

Did you ever had the experience that you where training a network, investing a lot of time in finding the right hyper parameters and testing different initializations to push that validation accuracy over certain threshold? Only to then find out when putting the model into production, that it significantly underperforms?

If you did, then you experienced one common problem with deep neural networks. The performance gap between in and out-of distribution generalization.

Today on the show PhD Rahim Entezari is giving us a wonderful tour through his PhD journey investigating ways to understand and improve generalization performance of deep neural networks.

Rahim will explain how one can improve generalization by different methods in data or in parameter space.

We will discuss how using different forms of sparsity, or the efficient creation of deep ensemble networks by permutation of network configurations can improve generalization from a parameter space perspective.

Or, from a data perspective where we discuss how data quality and data diversity effects the generalization performance of modern deep neural networks.

I hope you enjoy this interview, full of interesting concepts and ideas from deep learning theory.

# TOC

00:00:00 Introduction

00:02:18 Background Knowledge

00:06:56 Guest Introduction

00:12:35 Generalization from a Data or Parameter Perspective

00:16:21 In and out of distribution Generalization

00:20:30 Structured and Unstructured Sparsity

00:29:55 Generalization in Parameter space

00:46:56 Generalization in Data space

# Sponsors

Quantics: Supply Chain Planning for the new normal - the never normal - https://quantics.io/

Belichberg GmbH: We do digital transformations as your innovation partner - https://belichberg.com/

# References

Rahim Entezari - https://www.linkedin.com/in/rahimentezari/

Austrian Artificial Intelligence Podcast

Hosted by Manuel Pasieka, the Austrian Artificial Intelligence Podcast offers a grounded, local perspective on a global phenomenon. Instead of abstract theorizing, each conversation focuses on the tangible impact and practical applications of AI within Austria's unique ecosystem. You'll hear from a diverse range of guests-researchers, entrepreneurs, policymakers, and creatives-who are actively shaping this landscape, discussing both the remarkable opportunities and the nuanced challenges specific to the region. The discussions delve into how these technologies are being integrated into Austrian industry, academia, and society, moving beyond hype to examine real-world implementation and ethical considerations. This podcast serves as an essential audio forum for anyone in Austria, or with an interest in the European tech scene, looking to understand how artificial intelligence is evolving right here. It’s about the people behind the algorithms and the local stories within a global revolution. For those engaged with the content, questions and suggestions are always welcome at the provided email address.

Author: Manuel Pasieka Language: English Episodes: 73

Official website RSS

Austrian Artificial Intelligence Podcast

Podcast Episodes

[not-audio_url]

[/not-audio_url]

61. Jules Salzinger - AIT - Building explainable and generalizable AI Systems for Agriculture

03.10.2024

Duration: 1:26:20

Today on the podcast I have to pleasure to talk to Jules Salzinger, Computer Vision Researcher at the Vision & Automation Center of the AIT, the Austrian Institute of Technology. Jules will share with us, his newest rese…

[not-audio_url]

[/not-audio_url]

60. Alexandre Paris - Proofcheck - LLM fine-tuning and customization

28.08.2024

Duration: 53:19

## Summary Today on the show I am talking to Proofreads CTO Alexandre Paris. Alex explains in great detail how they analyze digital books drafts to identify mistakes and instances within the document that dont follow gui…

[not-audio_url]

[/not-audio_url]

59. Philip Winter - VRVis - Continual Learning

07.08.2024

Duration: 1:06:23

Today I am talking to Philip Winter, researcher at the Medical Imaging group of the VRVis, a research center for virtual realities and visualizations. Philip will explain the benefits and challenges in continual learning…

[not-audio_url]

[/not-audio_url]

58. Christa Zoufal - Quantum Machine Learning

16.07.2024

Duration: 59:43

## Summary AI is currently dominated by Deep Learning and Large Language Models, but there is other very interesting research that has the potential to have great impact on our lives in the future; one of them being Quan…

[not-audio_url]

[/not-audio_url]

57. Eldar Kurtic - Efficient Inference through sparsity and quantization - Part 2/2

25.06.2024

Duration: 46:38

Hello and welcome back to the AAIP This is the second part of my interview with Eldar Kurtic and his research on how to optimiz inference of deep neural networks. In the first part of the interview, we focused on sparsit…

[not-audio_url]

[/not-audio_url]

56. Eldar Kurtic - Efficient Inference through sparsity and quantization - Part 1/2

07.06.2024

Duration: 51:59

Hello and welcome back to the AAIP If you are an active Machine Learning engineer or are simply interested in Large Language models, I am sure you have seen the discussions around quantized models and all kind of new fra…

[not-audio_url]

[/not-audio_url]

55. Veronika Vishnevskaia - Ontec - Building RAG based Question-Answering Systems

15.04.2024

Duration: 1:10:48

## Summary Today on the show I am talking to Veronika Vishnevskaia. Solution Architect at ONTEC where she specialises in building RAG based Question-Answering systems. Veronika will provide a deep dive into all relevant…

[not-audio_url]

[/not-audio_url]

54. Manuel Reinsperger - MLSec & LLM Security

25.03.2024

Duration: 1:05:05

# Summary Today on the show I am talking to Manuel Reinsperger, Cybersecurity Expert and Penetration Tester. Manuel will provide us an introduction into the topic of Machine Learning Security with an emphasis on Chatbot…

[not-audio_url]

[/not-audio_url]

53. Peter Jeitscko - Impact of EU AI Regulation on AI startups

04.03.2024

Duration: 57:32

## Summary At the end of last year, the EU-AI Act was finalized and it spawned many discussions and a lot of doubts about the future of European AI companies. Today on the show Peter Jeitschko, founder of JetHire an AI b…

[not-audio_url]

[/not-audio_url]

52. Markus Keiblinger - Texterous - Building custom LLM Solutions

13.02.2024

Duration: 46:54

# Summary For the last two years AI has been flooded with news about LLMs and their successes, but how many companies are actually making use of them in their products and services? Today on the show I am talking to Mark…