Thoughtful discussions about current topics, moderated by American Banker editors.
…
continue reading
Deep Papers is a podcast series featuring deep dives on today’s most important AI papers and research. Hosted by Arize AI founders and engineers, each episode profiles the people and techniques behind cutting-edge breakthroughs in machine learning.
…
continue reading
Relevant, Inspirational, and Transformative. Explore Judaism with Rabbi Einhorn
…
continue reading
In this episode, we dive into the intriguing mechanics behind why chat experiences with models like GPT often start slow but then rapidly pick up speed. The key? The KV cache. This essential but under-discussed component enables the seamless and snappy interactions we expect from modern AI systems. Harrison Chu breaks down how the KV cache works, h…
…
continue reading
1
‘Job satisfaction will go up’: How generative AI is changing work
20:00
20:00
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
20:00
The technology will take on routine, dull work, says Alenka Grealish, principal analyst at Celent.
…
continue reading
In this byte-sized podcast, Harrison Chu, Director of Engineering at Arize, breaks down the Shrek Sampler. This innovative Entropy-Based Sampling technique--nicknamed the 'Shrek Sampler--is transforming LLMs. Harrison talks about how this method improves upon traditional sampling strategies by leveraging entropy and varentropy to produce more dynam…
…
continue reading
1
Google's NotebookLM and the Future of AI-Generated Audio
43:28
43:28
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
43:28
This week, Aman Khan and Harrison Chu explore NotebookLM’s unique features, including its ability to generate realistic-sounding podcast episodes from text (but this podcast is very real!). They dive into some technical underpinnings of the product, specifically the SoundStorm model used for generating high-quality audio, and how it leverages a hie…
…
continue reading
1
How banks’ use of generative AI has evolved over the past year
19:53
19:53
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
19:53
Financial institutions have dramatically increased their investment and trust in large language models since 2023. Kartik Ramakrishnan, Capgemini’s deputy CEO of Financial Services and head of banking and capital markets, shares the results of a recent report that analyzed these changes.
…
continue reading
1
Exploring OpenAI's o1-preview and o1-mini
42:02
42:02
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
42:02
OpenAI recently released its o1-preview, which they claim outperforms GPT-4o on a number of benchmarks. These models are designed to think more before answering and handle complex tasks better than their other models, especially science and math questions. We take a closer look at their latest crop of o1 models, and we also highlight some research …
…
continue reading
1
Upstart’s CEO Dave Girouard explains brighter outlook for rest of 2024
22:49
22:49
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
22:49
Advances in the company’s AI-based lending models have made them better at predicting risk, which has led to growth, he says.
…
continue reading
1
Breaking Down Reflection Tuning: Enhancing LLM Performance with Self-Learning
26:54
26:54
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
26:54
A recent announcement on X boasted a tuned model with pretty outstanding performance, and claimed these results were achieved through Reflection Tuning. However, people were unable to reproduce the results. We dive into some recent drama in the AI community as a jumping off point for a discussion about Reflection 70B. In 2023, there was a paper wri…
…
continue reading
1
Composable Interventions for Language Models
42:35
42:35
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
42:35
This week, we're excited to be joined by Kyle O'Brien, Applied Scientist at Microsoft, to discuss his most recent paper, Composable Interventions for Language Models. Kyle and his team present a new framework, composable interventions, that allows for the study of multiple interventions applied sequentially to the same language model. The discussio…
…
continue reading
1
‘These models will always hallucinate’: Seth Dobrin on LLMsDek
18:23
18:23
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
18:23
Dobrin, founder of advisory firm Qantm AI and former global chief AI officer at IBM, warns that popular generative AI models were trained on the whole of the internet and hallucinate at an unacceptable rate.
…
continue reading
1
Judging the Judges: Evaluating Alignment and Vulnerabilities in LLMs-as-Judges
39:05
39:05
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
39:05
This week’s paper presents a comprehensive study of the performance of various LLMs acting as judges. The researchers leverage TriviaQA as a benchmark for assessing objective knowledge reasoning of LLMs and evaluate them alongside human annotations which they find to have a high inter-annotator agreement. The study includes nine judge models and ni…
…
continue reading
1
‘Fraud is pervasive throughout the entire industry’: Crypto insider
21:34
21:34
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
21:34
Jake Donoghue, author of the book Crypto Confidential, shares some of the worst practices he saw as a founder of a cryptocurrency company.
…
continue reading
1
Breaking Down Meta's Llama 3 Herd of Models
44:40
44:40
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
44:40
Meta just released Llama 3.1 405B–according to them, it’s “the first openly available model that rivals the top AI models when it comes to state-of-the-art capabilities in general knowledge, steerability, math, tool use, and multilingual translation.” Will the latest Llama herd ignite new applications and modeling paradigms like synthetic data gene…
…
continue reading
1
Some banks are making a Faustian bargain with fintechs: Karen Petrou
18:57
18:57
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
18:57
Karen Petrou, the managing partner at Federal Financial Analytics and a long-time observer of banking and regulation, says banks need to do far more due diligence on potential fintech partners and exert more control over these relationships.
…
continue reading
1
DSPy Assertions: Computational Constraints for Self-Refining Language Model Pipelines
33:57
33:57
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
33:57
Chaining language model (LM) calls as composable modules is fueling a new way of programming, but ensuring LMs adhere to important constraints requires heuristic “prompt engineering.” The paper this week introduces LM Assertions, a programming construct for expressing computational constraints that LMs should satisfy. The researchers integrated the…
…
continue reading
1
What military members need from their banks
25:05
25:05
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
25:05
Two veterans and executives at Armed Forces Bank – Tom McLean and Jodi Vickery – share the challenges they see their customers face and new products the bank has rolled out this year to better serve them.
…
continue reading
1
Regulators are wise to be more careful’ after Chevron ruling
47:07
47:07
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
47:07
Gene Scalia, the banking lobby’s lawyer on retainer for a potential challenge to Washington’s capital reform effort, discusses the state of administrative law after the overturning of a key legal precedent.
…
continue reading
1
RAFT: Adapting Language Model to Domain Specific RAG
44:01
44:01
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
44:01
Where adapting LLMs to specialized domains is essential (e.g., recent news, enterprise private documents), we discuss a paper that asks how we adapt pre-trained LLMs for RAG in specialized domains. SallyAnn DeLucia is joined by Sai Kolasani, researcher at UC Berkeley’s RISE Lab (and Arize AI Intern), to talk about his work on RAFT: Adapting Languag…
…
continue reading
1
Climate First Bank's plans to expand nationwide
30:17
30:17
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
30:17
The Florida bank has emerged from de novo status and can now offer its solar loans in a larger geographic footprint.
…
continue reading
1
LLM Interpretability and Sparse Autoencoders: Research from OpenAI and Anthropic
44:00
44:00
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
44:00
It’s been an exciting couple weeks for GenAI! Join us as we discuss the latest research from OpenAI and Anthropic. We’re excited to chat about this significant step forward in understanding how LLMs work and the implications it has for deeper understanding of the neural activity of language models. We take a closer look at some recent research from…
…
continue reading
1
Can data ownership be preserved in generative AI?
16:46
16:46
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
16:46
Foundational models like GPT-4, the large language model behind ChatGPT, have hoovered up content from publications like The New York Times and social media sites like Reddit and OpenAI, and it faces several lawsuits because of this. John Thompson, global head of artificial intelligence at EY and author of the book Data for All, has set up what is …
…
continue reading
1
Trustworthy LLMs: A Survey and Guideline for Evaluating Large Language Models' Alignment
48:07
48:07
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
48:07
We break down the paper--Trustworthy LLMs: A Survey and Guideline for Evaluating Large Language Models' Alignment. Ensuring alignment (aka: making models behave in accordance with human intentions) has become a critical task before deploying LLMs in real-world applications. However, a major challenge faced by practitioners is the lack of clear guid…
…
continue reading
1
What might digital identity look like in the future?
21:30
21:30
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
21:30
Proof of identity is critical for many things, including being able to open a bank account, get a job, or obtain health care. Yet proving one’s identity is getting harder in a world of frequent data breaches. We asked Mariana Dahan, founder of the World Identity Network and chair of the Universal ID Council, what she thinks will solve this problem.…
…
continue reading
1
Breaking Down EvalGen: Who Validates the Validators?
44:47
44:47
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
44:47
Due to the cumbersome nature of human evaluation and limitations of code-based evaluation, Large Language Models (LLMs) are increasingly being used to assist humans in evaluating LLM outputs. Yet LLM-generated evaluators often inherit the problems of the LLMs they evaluate, requiring further human validation. This week’s paper explores EvalGen, a m…
…
continue reading
1
“The law was very clear” inside the Fed master account debate
36:53
36:53
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
36:53
Custodia Founder and CEO Caitlin Long says the Federal Reserve has rewritten the rules around accessing the government's payments system. The central bank and a federal court judge disagree. Editor’s note: This conversation was recorded on April 17. On April 26, Custodia Bank filed a notice of appeal, signaling that it will challenge the district c…
…
continue reading
1
Keys To Understanding ReAct: Synergizing Reasoning and Acting in Language Models
45:07
45:07
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
45:07
This week we explore ReAct, an approach that enhances the reasoning and decision-making capabilities of LLMs by combining step-by-step reasoning with the ability to take actions and gather information from external sources in a unified framework. To learn more about ML observability, join the Arize AI Slack community or get the latest on our Linked…
…
continue reading
1
'There are risks': Betsy Cohen on banking as a service
17:17
17:17
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
17:17
Banking as a service is expensive, it takes time and onboarding has to be done carefully, says the founder of Bancorp Bank, who now runs a venture capital firm that invests in fintechs.
…
continue reading
1
Is ‘technofeudalism’ killing capitalism?
25:51
25:51
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
25:51
Cloud-based tech giants like Amazon, Google and Uber are changing the economy, and not for the better, asserts Yanis Varoufakis, a former finance minister of Greece and a professor at the University of Athens, who has written a book about the dangers of what he calls the "cloudalists."
…
continue reading
1
Demystifying Chronos: Learning the Language of Time Series
44:40
44:40
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
44:40
This week, we’ve covering Amazon’s time series model: Chronos. Developing accurate machine-learning-based forecasting models has traditionally required substantial dataset-specific tuning and model customization. Chronos however, is built on a language model architecture and trained with billions of tokenized time series observations, enabling it t…
…
continue reading
1
‘Not all fintech is good for people’: Jennifer Tescher
25:45
25:45
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
25:45
Financial technology startups have developed some useful technology for consumers, such as automated savings, says Tescher, who founded the Financial Health Network 20 years ago. But some fintech innovations are more questionable.
…
continue reading
This week we dive into the latest buzz in the AI world – the arrival of Claude 3. Claude 3 is the newest family of models in the LLM space, and Opus Claude 3 ( Anthropic's "most intelligent" Claude model ) challenges the likes of GPT-4. The Claude 3 family of models, according to Anthropic "sets new industry benchmarks," and includes "three state-o…
…
continue reading
1
Reinforcement Learning in the Era of LLMs
44:49
44:49
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
44:49
We’re exploring Reinforcement Learning in the Era of LLMs this week with Claire Longo, Arize’s Head of Customer Success. Recent advancements in Large Language Models (LLMs) have garnered wide attention and led to successful products such as ChatGPT and GPT-4. Their proficiency in adhering to instructions and delivering harmless, helpful, and honest…
…
continue reading
1
How banks are helping the fight against illegal wildlife trading
26:17
26:17
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
26:17
Geraldine Fleming, financial task force manager at United for Wildlife and Jonny Bell, director, EMEA, LexisNexis Risk Solutions explain how banks around the world are helping to catch criminals who illegally mutilate, kill and sell rhinoceroses, elephants, donkeys and other animals.
…
continue reading
1
Sora: OpenAI’s Text-to-Video Generation Model
45:08
45:08
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
45:08
This week, we discuss the implications of Text-to-Video Generation and speculate as to the possibilities (and limitations) of this incredible technology with some hot takes. Dat Ngo, ML Solutions Engineer at Arize, is joined by community member and AI Engineer Vibhu Sapra to review OpenAI’s technical report on their Text-To-Video Generation Model: …
…
continue reading
1
‘Don’t fall for the sales pitches’: Advice on deploying AI
25:38
25:38
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
25:38
Eric Siegel, author of the book The AI Playbook, explains what it takes to take traditional and advanced artificial intelligence projects from idea to execution.
…
continue reading
This week, we’re discussing "RAG vs Fine-Tuning: Pipelines, Tradeoff, and a Case Study on Agriculture." This paper explores a pipeline for fine-tuning and RAG, and presents the tradeoffs of both for multiple popular LLMs, including Llama2-13B, GPT-3.5, and GPT-4. The authors propose a pipeline that consists of multiple stages, including extracting …
…
continue reading
We dive into Phi-2 and some of the major differences and use cases for a small language model (SLM) versus an LLM. With only 2.7 billion parameters, Phi-2 surpasses the performance of Mistral and Llama-2 models at 7B and 13B parameters on various aggregated benchmarks. Notably, it achieves better performance compared to 25x larger Llama-2-70B model…
…
continue reading
1
HyDE: Precise Zero-Shot Dense Retrieval without Relevance Labels
36:22
36:22
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
36:22
We discuss HyDE: a thrilling zero-shot learning technique that combines GPT-3’s language understanding with contrastive text encoders. HyDE revolutionizes information retrieval and grounding in real-world data by generating hypothetical documents from queries and retrieving similar real-world documents. It outperforms traditional unsupervised retri…
…
continue reading
1
How challenger bank Upgrade grew during a dismal 2023
29:04
29:04
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
29:04
The neobank doubled its membership to five million consumers and hired 200 people last year. Founder and CEO Renaud Laplanche explains how it fared during a time when many fintechs struggled.
…
continue reading
1
How generative AI could reshape financial services in 2024
17:02
17:02
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
17:02
Mike Abbott, global banking lead at Accenture, shared some of his predictions and opinions for the year ahead.
…
continue reading
1
Has the fintech movement lived up to its promise?
34:43
34:43
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
34:43
The fintech revolution has been more successful at working with banks than at trying to replace them, points out Gene Ludwig, former Comptroller of the Currency, chair of the Ludwig Institute for Shared Economic Prosperity, and co-founder of Canapi Ventures. Those with “must have” products will fare far better in 2024 than those with “nice to have”…
…
continue reading
1
A Deep Dive Into Generative's Newest Models: Gemini vs Mistral (Mixtral-8x7B)–Part I
47:50
47:50
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
47:50
For the last paper read of the year, Arize CPO & Co-Founder, Aparna Dhinakaran, is joined by a Dat Ngo (ML Solutions Architect) and Aman Khan (Product Manager) for an exploration of the new kids on the block: Gemini and Mixtral-8x7B. There's a lot to cover, so this week's paper read is Part I in a series about Mixtral and Gemini. In Part I, we prov…
…
continue reading
1
What to expect from VCs in 2024: Amy Nauiokas, Anthemis Group
28:06
28:06
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
28:06
It's been a rough year for fintechs and for the venture capital firms that fund them. Venture capital flows into financial technology companies dropped by 36% year over year to $6 billion in the third quarter of 2023. But Amy Nauiokas, founder and CEO of Anthemis Group, is optimistic about 2024.
…
continue reading
1
How to Prompt LLMs for Text-to-SQL: A Study in Zero-shot, Single-domain, and Cross-domain Settings
44:59
44:59
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
44:59
We’re thrilled to be joined by Shuaichen Chang, LLM researcher and the author of this week’s paper to discuss his findings. Shuaichen’s research investigates the impact of prompt constructions on the performance of large language models (LLMs) in the text-to-SQL task, particularly focusing on zero-shot, single-domain, and cross-domain settings. Shu…
…
continue reading
1
2023 was a rough year for bank regulators. What might 2024 bring?
17:08
17:08
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
17:08
Over the past year, the national bank regulators’ oversight of Silicon Valley Bank, Signature Bank, Silvergate Capital and other banks that failed has been criticized. Reports of a toxic workplace at the FDIC have come to light. And the OCC hired a Deputy Comptroller and overseer of fintech who had easily discoverable falsehoods on his resume. Mich…
…
continue reading
1
The Geometry of Truth: Emergent Linear Structure in LLM Representation of True/False Datasets
41:02
41:02
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
41:02
For this paper read, we’re joined by Samuel Marks, Postdoctoral Research Associate at Northeastern University, to discuss his paper, “The Geometry of Truth: Emergent Linear Structure in LLM Representation of True/False Datasets.” Samuel and his team curated high-quality datasets of true/false statements and used them to study in detail the structur…
…
continue reading
1
What fintechs think of the CFPB’s proposed data-sharing rule
30:50
30:50
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
30:50
Penny Lee, president and CEO of the Financial Technology Association and Steve Boms, executive director of FDATA NA, explain what their members like about the proposed regulation and what they would change.
…
continue reading
1
Towards Monosemanticity: Decomposing Language Models With Dictionary Learning
44:50
44:50
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
44:50
In this paper read, we discuss “Towards Monosemanticity: Decomposing Language Models Into Understandable Components,” a paper from Anthropic that addresses the challenge of understanding the inner workings of neural networks, drawing parallels with the complexity of human brain function. It explores the concept of “features,” (patterns of neuron ac…
…
continue reading
1
How community banks can use tech to stay relevant
20:09
20:09
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
20:09
Community banks sometimes feel that they lack the budget and staff to compete with larger banks and fintechs on things like mobile and online banking, virtual assistants and most recently generative AI. Jim Perry, senior strategist at Market Insights, suggests steps they can and should take to stay relevant technology wise.…
…
continue reading
1
What might the Sam Bankman-Fried trial mean for banks?
17:12
17:12
Later Afspelen
Later Afspelen
Lijsten
Vind ik leuk
Leuk
17:12
The case is not really about cryptocurrency but about fraud, points out Seoyoung Kim, department chair and associate professor of finance and business analytics at the Leavey School of Business at Santa Clara University. But regulators and lawmakers are watching and the outcome of the trial will have repercussions throughout finance.…
…
continue reading