AI

Artificial intelligence, machine learning, and everything LLM

1508 episodes Page 3 of 76

#5507: Inside ChatGPT's Hidden Pipeline: 8 Models Per Reply

One chat turn isn't one model call. It's moderation, PII filtering, a tone-based router, and activation classifiers — here's the documented graph.

ai-orchestrationai-memoryrag

#5506: AI Is Way More Than Chatbots

The word "AI" quietly became a synonym for "chatbot." Here's what the full model ecosystem actually contains.

large-language-modelsai-modelstaxonomy

#5505: Send a Bot to Your Next Sales Call

Daniel wants a bot that takes his sales meetings, asks hard questions, and ends the call when the pitch is spray-and-pray.

ai-agentsprompt-engineeringconversational-ai

#5503: Inside the Hidden Image Generation Pipeline

That one-click image generator is secretly a graph of many models. We reconstruct the hidden pipeline behind Gemini and ChatGPT.

image-generationsynthidprompt-engineering

#5487: Why Your Spreadsheet Mangles José's Name

A deep dive into character sets, from ASCII to UTF-8, and why José becomes "José" in your spreadsheet.

unicodebidirectional-textdata-integrity

#5485: Fine-Tuning Parakeet for Hebrew and Your Own Jargon

NVIDIA's Parakeet beats Whisper on Android — but can you teach it Hebrew, or just your own jargon? Two answers, one much happier.

fine-tuningspeech-recognitioncustom-asr

#5482: When AI Edits Your Words: The Off Switch Problem

A tool that works reliably — and still gets switched off. What over-editing studies reveal about why AI rewrites more than you asked.

prompt-engineeringai-agentscontext-window

#5481: Small Models, Big Guardrails: PII Detection in ChatGPT

A tiny 600M parameter model is quietly scanning your ChatGPT tool calls for PII. Here's how that class of guardian model actually works.

small-language-modelsprivacyai-security

#5480: Writing a Personal Instruction for ChatGPT

Daniel's personal ChatGPT prompt gets a line-by-line critique — and the case for shorter, sharper custom instructions.

prompt-engineeringcontext-windowai-memory

#5478: When AI Hands Out Your Phone Number

A chatbot gave out a stranger's real phone number. What happens when the model can't forget it?

privacytraining-dataai-ethics

#5472: Can Anyone Scan a QR Code From a Landing Plane?

A New York printing startup wants to tow a QR code behind a plane on the airport approach path. We rate the idea.

qr-codesaviationland-ownership

#5469: Open Source's New Map: India, Croatia, Ethiopia

One in three new GitHub developers now comes from a country outside 2020's top ten. What that shift means for open source.

open-sourcesoftware-developmentglobal-employment

#5467: Small Models, Big Schemas: When JSON Constraints Backfire

Small models plus strict JSON schemas should be a safe bet. A 15,000-generation study found the opposite.

small-language-modelsmodel-context-protocolai-agents

#5466: Hebrew Words Hidden in English Text

Daniel wants a classifier that spots Hebrew written in Latin letters — and it turns out nobody's built one.

linguisticsbidirectional-textautomatic-speech-recognition

#5465: Packaging AI Pipelines So They Actually Get Reused

Daniel's podcast pipeline works — so why can't he reuse it? Recipes, containers, and Hebrew code-switching TTS.

text-to-speechdockerdependency-management

#5464: Your Keyboard's Hidden Data Problem

Your keyboard knows your email address, your phrases, your habits — and you can't take any of it with you.

keyboard-layoutsandroidsideloading

#5463: What Happens When You Click "Connect" on an AI Plugin

You click accept, the tool runs, and a credential you never saw is doing the work. Here's who actually owns it.

model-context-protocolcybersecuritydigital-identity

#5461: TTS Can't Pronounce Hebrew Inside English

Your TTS reads Hebrew words with English phonetics. Here's why — and why the obvious fix doesn't work yet.

text-to-speechbidirectional-textlarge-language-models

#5460: Four Small Models, One Android Phone: Does It Actually Work?

A chained on-device dictation pipeline — VAD, ASR, cleanup — and why "it feels smooth" isn't the same as knowing it works.

on-device-asrquantizationvoice-to-text

#5459: Rebuilding the Podcast: Turn-Taking, Buttons, and Chatterbox

Daniel wants a push-to-interrupt button and a live voice loop. The research says the button is easy and the live part is a research project.

speech-to-speechhuman-computer-interactiontext-to-speech