Tokenless

Interactive discovery

Explore the topic map

Follow the connections between themes, people, and ideas across the Tokenless archive in an interactive topic modeling map.

Machine Learning

View All
Uncertainty-Guided Data Augmentation for Engineers | Deep Dive - Yongmin Kwon

Uncertainty-Guided Data Augmentation for Engineers | Deep Dive - Yongmin Kwon

This session details a data-efficient method for training engineering surrogate models by using uncertainty quantification (UQ) to guide geometric data augmentation. Instead of random deformations, the approach lets the deep ensemble model identify its own knowledge gaps (epistemic uncertainty), then uses Free-Form Deformation (FFD) to generate new shapes specifically in those uncertain regions. This ensures every expensive simulation run yields maximally informative data, significantly improving model accuracy for a fixed computational budget across domains like structural mechanics and aerodynamics.

Q-learning with Flow-Matching Policies

Q-learning with Flow-Matching Policies

This talk explores methods for optimizing expressive, multi-modal policies, such as those based on flow-matching, with off-policy reinforcement learning. The speaker presents two novel algorithms, FQ-RL and CAM, designed to overcome the instability of backpropagation through multi-step generative models, enabling effective online self-improvement and adaptation for robotic manipulation tasks.

Graph Neural Networks Explained: A Clear Guide to GNN Basics & Models

Graph Neural Networks Explained: A Clear Guide to GNN Basics & Models

An introduction to Graph Neural Networks (GNNs), covering fundamental concepts like nodes, edges, and embeddings. This post delves into the core message-passing mechanism and provides a detailed overview of key architectures including GCN, GraphSAGE, GAT, GIN, and Graph Transformers, explaining their unique approaches and mathematical formulations.

Artificial Intelligence

View All
⚡️Every product of the future will be a living system  — Ronak Malde, Trajectory.ai

⚡️Every product of the future will be a living system — Ronak Malde, Trajectory.ai

Ronuk Malde, CEO of Trajectory.ai, discusses his journey from building AI coding agents at Windsurf to his current focus on continual learning for enterprise AI. He shares insights on leveraging real-world user data, the unique challenges of model acquisition, and how Trajectory.ai's platform, powered by innovations like scaled SDPO and a novel training stack, enables dynamic, always-learning AI models for diverse industries from legal to finance.

6 Things to Know about AIE World's Fair 2026

6 Things to Know about AIE World's Fair 2026

Discover the AI Engineering World's Fair 2026, the largest iteration yet, offering an unparalleled deep dive into AI engineering with expanded tracks on auto research, GPU specialization, and new verticals like finance and healthcare. Highlights include an innovative expo experience, exclusive leadership initiatives like the "Token Billionaires Program," and unique side events fostering community, including "Posters on AI" where attendees can defend their tweets. This event is designed to be a curated hub for practical, cutting-edge insights and networking in the AI/ML professional landscape.

The data black hole at the center of AI

The data black hole at the center of AI

AI progress is fundamentally driven by vast amounts of data and compute, rather than improvements in sample efficiency, creating a stark contrast with human learning. This essay explores the "black hole of data" powering AIs, quantifies the massive sample-efficiency gap between humans and machines, counters common objections, and discusses the implications for white-collar automation and future AI research.

Technology

View All
3‑2‑1 Backup Rule Explained: Protect Your Data from Disaster

3‑2‑1 Backup Rule Explained: Protect Your Data from Disaster

Jeff Crume outlines essential data resiliency strategies, starting with the 3-2-1 backup rule—three copies, two media types, one offsite—and expanding to include immutable or air-gapped backups, rigorous testing, and encryption. He emphasizes these principles for robust disaster recovery, ransomware protection, and minimizing costly downtime, highlighting the trade-offs in achieving high availability.

The Media Game Has Changed

The Media Game Has Changed

The conversation explores the shift from legacy media to creator-led platforms, why authenticity has become a competitive advantage, and how founders can build audiences by communicating directly with customers, employees, and the public. They discuss podcasts, social media, storytelling, corporate communications, and the changing relationship between companies, journalists, and audiences. Along the way, they examine how founders can develop a public voice, why some leaders become influential communicators, and what it means to build a brand in a world where distribution is increasingly decentralized.

The C4 Model: Visualizing Software Architecture • Simon Brown & Susanne Kaiser • GOTO 2026

The C4 Model: Visualizing Software Architecture • Simon Brown & Susanne Kaiser • GOTO 2026

Simon Brown, creator of the C4 Model, discusses its origin as a practical solution to clarify messy software diagrams. He explains the four hierarchical levels (context, container, component, code), emphasizing that most teams only need the top two for significant value. The discussion highlights the importance of including technology in diagrams, C4's collaborative nature, and practical advice on modeling microservices and bounded contexts, all while advocating for a lightweight, accessible approach to architectural visualization.


Recent Post

What Breaks When You Build AI Under Sovereignty Constraints - Bilge Yücel, deepset GmbH

What Breaks When You Build AI Under Sovereignty Constraints - Bilge Yücel, deepset GmbH

Bilge Yücel explains that true AI sovereignty is a technical challenge, not just a policy issue, built on four pillars: data, model, infrastructure, and operations. This talk explores the common pitfalls of retrofitting sovereignty—from performance regressions after swapping models to discovering deep vendor lock-in when moving on-prem—and presents a checklist for building genuinely sovereign systems.

Personalization in the Era of LLMs - Shivam Verma, Spotify

Personalization in the Era of LLMs - Shivam Verma, Spotify

Spotify is personalizing open-weight LLMs without full fine-tuning by combining three key components: foundational user embeddings from streaming history, 'Semantic IDs' that tokenize its 100M+ item catalog, and a 'soft tokenization' layer that projects a user's embedding directly into the LLM's context. This allows the model to autoregressively generate the next song or podcast as the next token in a sequence.

How to Build Your Full-Stack Applications With CDK & AWS Amplify • Erik Hanchett • GOTO 2025

How to Build Your Full-Stack Applications With CDK & AWS Amplify • Erik Hanchett • GOTO 2025

Erik Hanchett from AWS demonstrates how to accelerate the development of full-stack applications using AWS Amplify Gen 2. The talk focuses on its TypeScript-first approach for end-to-end type safety, its abstraction over the AWS CDK, and the integration of generative AI features through the Amplify AI Kit to connect with services like Amazon Bedrock.

How to Build AI-First Organizations — with Jacob Miller and Jeremy Mumford

How to Build AI-First Organizations — with Jacob Miller and Jeremy Mumford

Jacob Miller and Jeremy Mumford, authors of 'Architected Intelligence', discuss the enduring principles for building successful AI products and organizations. They cover why velocity is the only durable moat, why hallucinations are a data curation issue, and the proper progression from skills to workflows to agents, emphasizing a shift from focusing on models to focusing on process and speed.

Let's go Bananas with GenMedia — Guillaume Vernade, Google DeepMind

Let's go Bananas with GenMedia — Guillaume Vernade, Google DeepMind

Guillaume Vernade from Google DeepMind demonstrates a full generative media pipeline, using Gemini to read a public domain book and act as a master prompt engineer for other models. Imagen generates character portraits, Veo animates scenes into video, Lyria composes a unique soundtrack for each chapter, and a clever TTS trick creates a multi-character audiobook.

Build Agents That Run for Hours (Without Losing the Plot) — Ash Prabaker & Andrew Wilson, Anthropic

Build Agents That Run for Hours (Without Losing the Plot) — Ash Prabaker & Andrew Wilson, Anthropic

Explore advanced techniques for building long-running AI agents, moving beyond simple loops. Learn why self-evaluation fails and adversarial evaluators succeed, how to manage context with structured handoffs instead of just compaction, and how to use negotiated 'sprint contracts' and detailed rubrics to build and test complex, full-stack applications autonomously.

Stay In The Loop! Subscribe to Our Newsletter.

Get updates straight to your inbox. No spam, just useful content.