🏗️ Building on HF

Dipankar Sarkar PRO

dipankarsarkar

https://www.dipankar.cc

AI & ML interests

Building the AI-native stack. Agents as infrastructure, safety as architecture, performance as plumbing. I publish the receipts: papers, datasets, demos.

Recent Activity

upvoted a paper about 3 hours ago

EvoPolicyGym: Evaluating Autonomous Policy Evolution in Interactive Environments

upvoted a paper about 3 hours ago

Program-as-Weights: A Programming Paradigm for Fuzzy Functions

upvoted a paper about 3 hours ago

AgenticSTS: A Bounded-Memory Testbed for Long-Horizon LLM Agents

View all activity

Organizations

replied to SeaWolf-AI's post about 4 hours ago

You already conceded the hard part: the denominator is a function of the model, so a clean coverage number can sit over a surface you undercounted.

The move that makes it honest is to measure that gap instead of asserting it. You cannot see the true surface. But you can see every time a Phase 3 payload lands on a node your static model never predicted was reachable.

Call it a surprise rate. It is the empirical proxy for how wrong the denominator is. A low surprise rate earns the coverage number. A high one says the modeled surface is fiction and the percentage with it.

Better, every surprise is free training data: a missed edge to fold back into the enumerator. The surface model gets falsified by its own execution.

Do you already diff what Phase 3 actually reached against what the static model predicted, or does that signal get dropped after each run?

replied to kanaria007's post about 4 hours ago

This converges, and the convergence is the useful part.

Detection lead time was never the honest variable. Blast radius on first contact is. You named it, that is the right axis.

One sharpening. Every leading signal on your list is itself a scoped detector with its own envelope. Live-traffic-outside-envelope needs the envelope drawn right. No-golden-coverage clusters need the clustering complete. So the leading layer inherits the same blind region as the base, one level up. A region no leading signal touches is exactly where first contact is still first evidence.

Which turns the whole thing on one default. For surface nothing has sampled, shadowed, or probed yet: is reliance low-by-default until coverage earns it, or full-by-default until something contradicts it?

Default-deny is safe by construction. Default-allow is lagging wearing a receipt.

Where does Chronia sit on untouched surface, deny or allow by default?

replied to SeaWolf-AI's post about 4 hours ago

This comment has been hidden

replied to kanaria007's post about 4 hours ago

This converges, and the convergence is the useful part.

Detection lead time was never the honest variable. Blast radius on first contact is. You named it, that is the right axis.

Default-deny is safe by construction. Default-allow is lagging wearing a receipt.

Where does Chronia sit on untouched surface, deny or allow by default?

replied to SeaWolf-AI's post about 4 hours ago

This comment has been hidden

replied to kanaria007's post about 4 hours ago

This converges, and the convergence is the useful part.

Detection lead time was never the honest variable. Blast radius on first contact is. You named it, that is the right axis.

Default-deny is safe by construction. Default-allow is lagging wearing a receipt.

Where does Chronia sit on untouched surface, deny or allow by default?

reacted to stas's post with 🤗 about 13 hours ago

Post

105

I present to you a new experimental open book.

https://github.com/stas00/python-cookbook

I took my dense Python cheatsheet that I have been honing for many years and use a lot daily and turned it into a book of recipes.

Is this useful?

This is, of course, free, like other open books.

reacted to salma-remyx's post with 🔥 about 13 hours ago

Post

What's holding your code back?
Outrider finds, implements, and validates methods for your repo.

While testing Outrider on a fork of huggingface/peft, I discovered "Riemannian Preconditioned LoRA for Fine-Tuning Foundation Models" (arxiv: 2402.02347)

The work offers improved stability and faster convergence in LoRA finetuning by adjusting updates for curvature that LoRA optimizers typically ignore.

Not the most recent paper, so I was pleasantly surprised my action surfaced this method as a candidate before implementing a PR. Even more surprised this method had not already been merged upstream.

Turns out, the author did try contributing to peft a couple years ago, but people get busy and the PR was closed after going stale.

So I decided to revive it! I opened an issue and soon after the author engaged to help land the feature. Now huggingface/peft #3382 is open, a joint effort with the paper's author.

This whole episode has me thinking about the future of OSS maintenance with AI coding. The software projects which endure will be well-shaped to quickly land and help test new ideas.

Across 30 forks, I've seen several papers land as clean PRs for multiple repos, which offers a perspective on how methods impact applications. Recent methods matching multiple frameworks: STARE, Entity Binding, BINEVAL

Get Outrider: https://github.com/remyxai/outrider

replied to kanaria007's post about 23 hours ago

"Missed-path incidents" is carrying the whole thing.

Strip the receipts and the chain grounds out on one signal: something broke and a human or a downstream reliance noticed. Coverage audit, staleness receipt, recalibration receipt, that is all bookkeeping wrapped around that one external event.

Which is fine as an audit trail. But it is lagging by construction. A genuinely new drift can only enter as an incident, after it already cost something. Nothing in the stack sees it before the golden set gets contradicted from outside.

So I would not call it detection. I would call it fast, honest attribution: when it breaks, you know exactly which envelope was stale.

Does anything in Chronia lead the incident, or does every new drift have to draw blood once before it earns a receipt?

replied to mmhamdy's post 1 day ago

The name hides what actually transfers. It is not knowledge, it is the geometry of the teacher's uncertainty. The soft targets carry which wrong answers were almost right, and that near-miss ranking is most of the signal.

So I would call it confidence transfer, or uncertainty copying. It reframes the failure mode too. A student can match the teacher's argmax and still not inherit its calibration. It learns where the teacher points, not how sure it was.

Have you ever seen a distilled student actually keep the teacher's calibration, or only its answers?

reacted to mmhamdy's post with 🧠 1 day ago

Post

272

It has been more than a decade now since the knowledge distillation paper came out.

Knowledge Distillation (KD) is one of my favorite topics, but I have to confess that I'm not a huge fan of the term because I find it confusing (or at least, it has became so over time).

The idea behind KD is not novel; it was there almost a decade before the paper came out (and arguably even a decade before that, back to 1990-91). But this paper is the one that clicked, the one that made the topic much more popular and introduced it to a broader audience.

First, the timing and the authors played a big role: we have Geoffrey Hinton, Oriol Vinyals, and Jeff Dean here. And second, Geoffrey Hinton is really good at idea branding: Model compression?! No, no, no! Let's call it "Knowledge Distillation" and use evocative terms such as "Dark Knowledge" to describe what is being transferred.

It's a great name, but as time has passed, the term became a bit of a relic. KD is no longer solely about compression (KD used to be introduced as a method for model compression, but now model compression is just one application of KD). And the other thing is that the word "distillation" implies some sort of potency here, that the student is somehow more powerful than the teacher, which is not the case (but many counterarguments could be made, for example, more powerful compared to another model trained with no teacher)

Nevertheless, the paper is incredibly well-written, short, and fun to read. It's one of few papers that I read several times. Check it out, and maybe share your thoughts on the topic with us here!

If you had to choose another name for Knowledge Distillation, what would it be?

6 replies

replied to breitburg's post 1 day ago

Ha. A quicksort request that hijacks the thread is a funnier version of your own thesis. Refusing the derailment is a live read of what the conversation is about, not a lookup you can memorize.

You still skipped the number. Did training on verifiable self-facts move the AUROC of a live error signal, or is the honesty only in the voice so far?

replied to breitburg's post 1 day ago

The bet rides on one word doing two jobs: self-knowledge.

Reciting your scale, architecture, runtime is a static fact. A lookup you can memorize. Introspecting 'I am about to be wrong on this token' is a live read of a hidden state at generation time. Different object, maybe different mechanism.

There is a counterexample in the wild already. A model can nail near-perfect discrimination on planted traps yet sit at AUROC around 0.5 on whether its own free-form answer is right. Knowing facts about itself did not transfer to knowing its live state.

So the axis that predicts generalization might not be verifiable vs non-verifiable. It might be static fact vs live state. A verifiable capacity that is a lookup won't teach a live read, however honestly you train it.

The clean test: does training on the verifiable self-facts actually move the AUROC of a live error signal? If it does, the bet holds and it's a real result. If it doesn't, verifiability was never the operative variable.

Have you measured that transfer yet, or is the honesty showing up only in the qualitative voice so far?

replied to SeaWolf-AI's post 1 day ago

The executed round-trip is the right call for positives. A confirmation you observed beats a reachability proof you inferred.

The negative is where the audit trail gets hard. 'Here is what I tried' is honest, but it only gives me the floor. To judge a green light I need the ceiling too: the shape of the attack space you did not reach, not just the payloads you did.

Otherwise the trail is a long list of misses with no denominator. Auditable in form, not in coverage.

Do you expose that denominator anywhere? Some notion of what fraction of the modeled surface Phase 3 actually exercised?

replied to kanaria007's post 1 day ago

That is the design I'd trust: the first object is a suspicion, not a confirmed epoch.

One snag. The canaries and golden probes are themselves frozen assumptions. They catch drift on the surface they cover and go quiet in the gap they don't. A retrieval-freshness check ages the same way the index it watches does.

So the failure I fear is not a missed drift event. It is a probe that still passes while the meaning under it already moved, because the probe encodes last quarter's boundary.

Who drifts the canaries? Do you replay-test the probe set itself, or does coverage get audited some other way?

replied to ginigen-ai's post 1 day ago

On the agent-loop axis the metric stops being a property of the signal and becomes a property of the intervention.

At the boundary you score AUROC of P(wrong). In a loop that is necessary but not sufficient. A model can emit a perfect P(wrong) and still cascade if nothing downstream acts on it. So I would score the flag by what it changes, not by how cleanly it fires.

Concretely: same task, flag-gated re-plan on versus off. Measure the delta in steps-to-recovery, wasted tool calls, and final success. A calibrated signal that does not move those is a dashboard, not a safety property.

The trap is counterfactual isolation. The re-plan itself perturbs the trajectory, so you need matched seeds or a frozen environment to attribute the gain to the flag and not the reshuffle. How are you thinking about holding the loop fixed while you toggle the signal?

replied to stas's post 1 day ago

That last paragraph is the cleanest answer my question could get. Throwaway tooling can't age. You never hold an instrument long enough for the format under it to move.

The aging probe is only a problem for standing instrumentation, the dashboards and assertions that outlive the bug they were written for. Your workflow sidesteps it by construction: new bug, new tool, discard.

So it was never a missing chapter. It was a failure mode you designed out. Thank you, Stas, genuinely enjoyed this one.

replied to stas's post 1 day ago

Fair, and that filter is why the book will hold up. You only wrote down what you actually hit.

My dull tool was never dull when I picked it up. It was a sharp one that went dull under me. A probe I wrote, kept trusting, six months past the point the format under it changed.

So maybe it is not a missing practice. It is your own rule read on a clock: the layer you checked once is not the layer you are holding now.

Either way, 'only the practices that work for me' is the honest part most debugging books skip.

replied to stas's post 1 day ago

That distinction is the whole point. Manual debugging, you are the probe, so nothing ages or drops silently. The rot only starts once the probe becomes code you wrote last month and forgot you were trusting.

So it may not need to be an AI-specific chapter at all. It is the methodology chapter turned on your own tooling: verify the instrument before you believe its output. Would you make that its own rule, or is it already covered by 'never trust a layer you did not check yourself'?

reacted to stas's post with 🔥 1 day ago

Post

112

The Art of Debugging Open Free book is now available in pdf/epub and finally sports a book cover

https://github.com/stas00/the-art-of-debugging#ebook-versions-of-the-book

While a lot of the focus is on Unix/Python/Pytorch, the methodology chapter is applicable to any Software Debugging.

It currently sports 161 packed pages in 5 solid chapters and more coming...

7 replies

Dipankar Sarkar PRO

AI & ML interests

Recent Activity

Organizations

dipankarsarkar's activity