Inside the Black Box: Cracking AI and Deep Learning
Inside the Black Box: Cracking AI and Deep Learning

Inside the Black Box: Cracking AI and Deep Learning

How do Large Language Models like ChatGPT work, anyway?

Episodes (36)

Three Ways to Be Wrong About the Past Tense

Three Ways to Be Wrong About the Past Tense

Not at This Address

Not at This Address

My Advisors Argued This for Thirty Years. Now You Can Check

My Advisors Argued This for Thirty Years. Now You Can Check

Inside the Past Tense: Raw Letters, Raw Sound

Inside the Past Tense: Raw Letters, Raw Sound

Forcing a Neural Network to Add -ed

Forcing a Neural Network to Add -ed

The Wug Test for AI

The Wug Test for AI

Learning the Past Tense in AI

Learning the Past Tense in AI

Fine Tuning Lora: It's Not What You Think

Fine Tuning Lora: It's Not What You Think

When Fluent Answers Start Sounding True

When Fluent Answers Start Sounding True

Why Your Brain Believes the Model

Why Your Brain Believes the Model

When Polished Answers Feel Finished

When Polished Answers Feel Finished

What Seneca Teaches Us that Marcus Couldn't

What Seneca Teaches Us that Marcus Couldn't

The Pattern Holds for Another Author

The Pattern Holds for Another Author

The Pattern Holds

The Pattern Holds

Cracking Open the Black Box

Cracking Open the Black Box

Inside a Fine-Tuned Language Model

Inside a Fine-Tuned Language Model

What Counts as Structure? From Harris and Elman to Today’s Neural Nets

What Counts as Structure? From Harris and Elman to Today’s Neural Nets

Building a House Without Blueprints: When Interpretability Tools Work — and When They Don’t

Building a House Without Blueprints: When Interpretability Tools Work — and When They Don’t

I Told My LLM Not to Say "Empower"

I Told My LLM Not to Say "Empower"

Beyond the Surface of AI Intelligence

Beyond the Surface of AI Intelligence

Unlocking BERTs Hidden Grammar

Unlocking BERTs Hidden Grammar

Cracking the Code of AI Interpretation

Cracking the Code of AI Interpretation

Decoding GPTs Hidden Circuits

Decoding GPTs Hidden Circuits

Decoding Attention and Emergence in AI

Decoding Attention and Emergence in AI

When Knowledge Battles Noise in GPT Models

When Knowledge Battles Noise in GPT Models

Inside Circuits: How Large Language Models Understand

Inside Circuits: How Large Language Models Understand

Hallucinations, Interpretability, and the Seahorse Mirage

Hallucinations, Interpretability, and the Seahorse Mirage

How Transformers Stack Meaning Like Finnish Words

How Transformers Stack Meaning Like Finnish Words

The Mandela Effect in AI: Why Language Models Misremember

The Mandela Effect in AI: Why Language Models Misremember

Bridging Circuits and Concepts in Large Language Models

Bridging Circuits and Concepts in Large Language Models

How Transformers Turn Words Into Meaning

How Transformers Turn Words Into Meaning

Can Smaller Language Models Be Smarter?

Can Smaller Language Models Be Smarter?

The Weird Geometry That Makes AI Think

The Weird Geometry That Makes AI Think

Can We Fix It?

Can We Fix It?

Using Symbolic AI to Explain LLMs

Using Symbolic AI to Explain LLMs

Peering Inside the Black Box

Peering Inside the Black Box