<?xml version="1.0" encoding="utf-8" standalone="yes"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:content="http://purl.org/rss/1.0/modules/content/">
  <channel>
    <title>socratic-method on S Anand</title>
    <link>https://www.s-anand.net/blog/tag/socratic-method/</link>
    <description>Recent content in socratic-method on S Anand</description>
    <generator>Hugo -- 0.164.0</generator>
    <language>en-us</language>
    <lastBuildDate>Fri, 12 Jun 2026 08:10:56 +0530</lastBuildDate>
    <atom:link href="https://www.s-anand.net/blog/tag/socratic-method/index.xml" rel="self" type="application/rss+xml" />
    <item>
      <title>Let AI take your exams</title>
      <link>https://www.s-anand.net/blog/let-ai-take-your-exams/</link>
      <pubDate>Fri, 12 Jun 2026 08:10:56 +0530</pubDate>
      <guid>https://www.s-anand.net/blog/let-ai-take-your-exams/</guid>
      <description>&lt;p&gt;At 2 pm IST today (Fri 12 Jun 2026), I conducted a workshop at &lt;a href=&#34;https://www.iitmparadox.org/workshops&#34;&gt;Paradox, IITM&lt;/a&gt; - at &lt;a href=&#34;https://doms.iitm.ac.in/&#34;&gt;DOMS 101&lt;/a&gt;.&lt;/p&gt;
&lt;p&gt;My core message is: &amp;ldquo;AI can solve exams and help you learn. Delegate what AI can do. Learn what AI &lt;em&gt;can&amp;rsquo;t&lt;/em&gt; do instead.&amp;rdquo;&lt;/p&gt;
&lt;video controls preload=&#34;metadata&#34; width=&#34;1920&#34; height=&#34;1080&#34; style=&#34;max-width: 100%; height: auto;&#34;&gt;
  &lt;source src=&#34;https://media.s-anand.net/2026-06-12-let-ai-take-your-exam.webm&#34; type=&#34;video/webm; codecs=&amp;quot;vp9, opus&amp;quot;&#34;&gt;
&lt;/video&gt;
&lt;p&gt;&lt;a href=&#34;https://sanand0.github.io/talks/2026-06-12-let-ai-take-your-exams/&#34;&gt;My talks page for &amp;ldquo;Let AI take your exams&amp;rdquo;&lt;/a&gt; includes:&lt;/p&gt;
&lt;ul&gt;
&lt;li&gt;&lt;a href=&#34;https://sanand0.github.io/talks/2026-06-12-let-ai-take-your-exams/story.html&#34;&gt;The full story + transcript + audio&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://sanand0.github.io/talks/2026-06-12-let-ai-take-your-exams/codex.html&#34;&gt;How Codex solved a real exam, live&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://sanand0.github.io/talks/2026-06-12-let-ai-take-your-exams/techniques.html&#34;&gt;My collection of AI-learning techniques&lt;/a&gt; - which was not covered in the workshop, but is a useful reference&lt;/li&gt;
&lt;/ul&gt;
&lt;p&gt;Here are the &lt;a href=&#34;https://sanand0.github.io/talks/2026-06-12-let-ai-take-your-exams/story.html&#34;&gt;takeaways from the workshop&lt;/a&gt;:&lt;/p&gt;
&lt;ol&gt;
&lt;li&gt;&lt;strong&gt;AI is more capable than you think — and getting smarter.&lt;/strong&gt; Recalibrate constantly what it can and can&amp;rsquo;t do. Note down what it can&amp;rsquo;t, because that is precisely where your value lives.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Delegate first; learn the rest.&lt;/strong&gt; Give everything to AI. Focus your learning on what it can&amp;rsquo;t yet do — that&amp;rsquo;s where the value will be. It&amp;rsquo;s a moving filter; revisit it every quarter.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Always use the best model, turned up high.&lt;/strong&gt; Reserve &amp;ldquo;fast and cheap&amp;rdquo; for the ~5% of moments you need a quick answer. And remember students get serious AI free via the GitHub Student Pack and Gemini.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Make the AI ask &lt;em&gt;you&lt;/em&gt; for context.&lt;/strong&gt; &amp;ldquo;If you need more information, ask me.&amp;rdquo; You don&amp;rsquo;t have to know what context it needs — push that burden back to the machine.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Beat hallucination with a maker-checker.&lt;/strong&gt; Two independent models that must agree cut errors from 14% to under 4%. Tell the checker to &amp;ldquo;find the errors,&amp;rdquo; not to grade.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Loop with feedback in verifiable environments.&lt;/strong&gt; Point an agent at an exam, a codebase, anything that scores itself — let it try, submit, read the result, retry. This is the most powerful technique AI has.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Calibrate, don&amp;rsquo;t just trust.&lt;/strong&gt; Practise predicting whether AI will get something right — even on topics you don&amp;rsquo;t know. Watch for base-rate traps and familiar problems with one changed premise.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Be lazy, productively.&lt;/strong&gt; Don&amp;rsquo;t read AI&amp;rsquo;s 20-page output — train it to give you five words. Working well with AI is a management skill.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Learn from peers.&lt;/strong&gt; Multiple people trying things is how you discover what works. Non-transactional relationships are the rising currency of the AI era.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Apply the scientific method to everything.&lt;/strong&gt; Form a hypothesis, hunt for evidence, try to falsify yourself. And when a system blocks you unfairly — hack it, then publish what you learned.&lt;/li&gt;
&lt;/ol&gt;
&lt;p&gt;Here are the &lt;a href=&#34;https://sanand0.github.io/talks/2026-06-12-let-ai-take-your-exams/codex.html&#34;&gt;takeaways from how Codex solved the exam&lt;/a&gt;:&lt;/p&gt;
&lt;ol&gt;
&lt;li&gt;&lt;strong&gt;An agent operates the environment; a chatbot answers the question.&lt;/strong&gt; Codex read source, ran code, clicked Check, and looped on feedback. That&amp;rsquo;s why it beat copy-paste — the exam was full of affordances a chatbot can&amp;rsquo;t touch.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Verifiable environments favour AI.&lt;/strong&gt; The more checkable the exam — validators, error strings, downloadable files, a live Check button — the more it helped the agent, not the student. &amp;ldquo;AI-proof&amp;rdquo; and &amp;ldquo;feedback-rich&amp;rdquo; are opposites.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Most failures were wording, not reasoning — and the source fixed them.&lt;/strong&gt; The fix for a brittle validator was to read the validator. Pass the error message back to the agent and it converges.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Don&amp;rsquo;t guess where attempts are limited.&lt;/strong&gt; The network game punished early guesses. The recoverable mistakes had feedback; the costly ones didn&amp;rsquo;t. Triage cheap-and-certain first.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;The gap between 9 and 10 was a credential, not a brain.&lt;/strong&gt; Same model, same skill. Anand&amp;rsquo;s missing mark was an invalid token. In the AI era, &amp;ldquo;can it?&amp;rdquo; often means &amp;ldquo;does it have the keys?&amp;rdquo;&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;It cost about a coffee.&lt;/strong&gt; ~$2–3 of tokens for the whole exam, ~96% cached. The real cost is the human judgment to know when the agent is plausibly, confidently wrong.&lt;/li&gt;
&lt;/ol&gt;
&lt;h3 id=&#34;original-announcement&#34;&gt;Original announcement&lt;/h3&gt;
&lt;p&gt;You can join online at &lt;a href=&#34;https://meet.google.com/cpt-faee-ucx&#34;&gt;https://meet.google.com/cpt-faee-ucx&lt;/a&gt; and ask questions on chat.&lt;/p&gt;
&lt;p&gt;&lt;strong&gt;Agenda&lt;/strong&gt;:&lt;/p&gt;
&lt;blockquote&gt;
&lt;p&gt;You&amp;rsquo;ve been told AI can pass your exams. But what happens when you actually watch it try — live, on your questions, in real time?&lt;/p&gt;
&lt;p&gt;This workshop starts with a collective experiment: we ask coding agents to solve real exams (including IITM exams) and see how it solves them.&lt;/p&gt;
&lt;p&gt;What follows isn&amp;rsquo;t a tutorial on prompting — it&amp;rsquo;s an autopsy that reveals what your exams are actually testing, where AI confidently hallucinates, and what that means for what&amp;rsquo;s worth learning.&lt;/p&gt;
&lt;p&gt;You&amp;rsquo;ll leave with a reframed understanding of your degree (the goal isn&amp;rsquo;t answers, it&amp;rsquo;s the ability to catch wrong ones) and a concrete study rituals that uses AI as a Socratic sparring partner rather than an answer machine.&lt;/p&gt;
&lt;p&gt;Come with a question you got wrong recently — it&amp;rsquo;s going to be useful.&lt;/p&gt;
&lt;/blockquote&gt;
&lt;p&gt;&lt;strong&gt;Real agenda&lt;/strong&gt;: An &lt;a href=&#34;https://en.wikipedia.org/wiki/R/IAmA&#34;&gt;ask-me-anything&lt;/a&gt; session plus real-life experiments.&lt;/p&gt;
&lt;p&gt;&lt;img loading=&#34;lazy&#34; src=&#34;https://files.s-anand.net/images/2026-05-18-let-ai-take-your-exams.avif&#34;&gt;&lt;/p&gt;
</description>
    </item>
    <item>
      <title>Things I Learned - 01 Feb 2026</title>
      <link>https://www.s-anand.net/blog/things-i-learned-01-feb-2026/</link>
      <pubDate>Sun, 01 Feb 2026 00:00:00 +0000</pubDate>
      <guid>https://www.s-anand.net/blog/things-i-learned-01-feb-2026/</guid>
      <description>&lt;p&gt;This week, I learned:&lt;/p&gt;
&lt;ul&gt;
&lt;li&gt;Android screen recorder is the easiest way to record phone and WhatsApp calls. But that won&amp;rsquo;t work for Google Meet, Teams, Zoom, etc. &lt;a href=&#34;https://gemini.google.com/share/1aee78cf4e15&#34;&gt;Gemini&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://exiftool.org/&#34;&gt;exiftool&lt;/a&gt; remains the best media metadata extractor (music, images, &amp;hellip;) though it&amp;rsquo;s old, slow, and Perl-based. &lt;code&gt;exiftool -csv -r ~/Music/ &amp;gt; music.csv&lt;/code&gt; exports all metadata as CSV. Installing the source via &lt;a href=&#34;https://sourceforge.net/projects/exiftool/files/latest/download&#34;&gt;https://sourceforge.net/projects/exiftool/files/latest/download&lt;/a&gt; seems best. It&amp;rsquo;s a good alternative to mp3tag / puddletag UI-based exports. &lt;a href=&#34;https://chatgpt.com/share/697c398a-5dac-8003-93a2-fabcd15d6b2e&#34;&gt;ChatGPT&lt;/a&gt; &lt;a href=&#34;https://gemini.google.com/share/4e95924dfac4&#34;&gt;Gemini&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;⭐ Some questions are for us to learn. Some are Socratic, and meant for the answerer to learn. When working with AI agents and interns, I find myself asking them several questions that I don&amp;rsquo;t want to know the answer for, but is important for them along their journey. Roughly the equivalent of &amp;ldquo;Think step by step&amp;rdquo; converted into the Socratic method. For example:
&lt;ul&gt;
&lt;li&gt;Instead of &amp;ldquo;Build a demo for this client&amp;rdquo;, ask &amp;ldquo;Who is the audience? What&amp;rsquo;s their objective?&amp;rdquo; and THEN ask for a demo.&lt;/li&gt;
&lt;li&gt;Instead of &amp;ldquo;Generate a dummy dataset for X&amp;rdquo;, ask &amp;ldquo;What interesting insights would we want when analyzing X?&amp;rdquo; and THEN ask for a dataset.&lt;/li&gt;
&lt;li&gt;Instead of &amp;ldquo;Write this code&amp;rdquo;, ask &amp;ldquo;What&amp;rsquo;s the best architecture for this?&amp;rdquo; and THEN ask for code.&lt;/li&gt;
&lt;/ul&gt;
&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://news.ycombinator.com/item?id=46549444&#34;&gt;Executable Markdown files with Unix pipes&lt;/a&gt; sounds like a clever idea. Prefix Markdown files with &lt;code&gt;#!/usr/bin/env codex&lt;/code&gt; (or &lt;code&gt;claude -p&lt;/code&gt;). Then, just write programs by describing them.&lt;/li&gt;
&lt;li&gt;Quotes from &lt;a href=&#34;https://www.goodreads.com/book/show/210300489-isles-of-the-emberdark&#34;&gt;Isles of the Emberdark&lt;/a&gt;:
&lt;ul&gt;
&lt;li&gt;Really, he should have known better than to punch a senator. Important people had underlings you punched on their behalf, and he should have found one of those.&lt;/li&gt;
&lt;/ul&gt;
&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://help.openai.com/en/articles/9930697-what-is-the-canvas-feature-in-chatgpt-and-how-do-i-use-it&#34;&gt;ChatGPT Canvas&lt;/a&gt; has a cool feature for editing documents or code. Just select a portion, ask for changes, and it edits it. Importantly, it&amp;rsquo;s very fast.&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://www.nationalgeographic.com/podcasts/greeking-out&#34;&gt;Greeking Out&lt;/a&gt; is a kid-friendly National Geographic podcast about ancient Greece and its influence on modern life.&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://fly.io/blog/code-and-let-live/&#34;&gt;fly.io containers&lt;/a&gt; at &lt;a href=&#34;https://sprites.dev/&#34;&gt;sprites.dev&lt;/a&gt; seem impressive. You can SSH into them. They have public &amp;amp; private HTTPS URLs. It auto-sleeps after 30s. You can checkpoint any time and restore the ENTIRE system. It&amp;rsquo;s FAST! This is great for agents. Just install Claude Code / Codex and other tools. Checkpoint it. Then &lt;code&gt;ssh&lt;/code&gt; into it and use as required. The cost is typically ~12c/hour - which is expensive to run forever but great for bursts. &lt;a href=&#34;https://simonwillison.net/2026/Jan/9/sprites-dev/&#34;&gt;Simon Willison&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;I&amp;rsquo;m seeing the Collider Bias in action (on a small sample). The developers who can communicate well don&amp;rsquo;t code as well, and vice versa. Not because there&amp;rsquo;s a negative correlation - but because I&amp;rsquo;m eliminating people who can &lt;em&gt;neither&lt;/em&gt; code nor communicate. But interestingly, over a 1-3 month horizon, the ones who code start communicating much better but the ones who communicate well don&amp;rsquo;t start coding much better. My theory is that the developers I work are communication-bottlenecked (e.g. lack of confidence) than unskilled (e.g. poor communicators).&lt;/li&gt;
&lt;li&gt;Prefer &lt;a href=&#34;https://zod.dev/&#34;&gt;Zod&lt;/a&gt; for TypeScript validation and &lt;a href=&#34;https://ajv.js.org/&#34;&gt;Ajv&lt;/a&gt; for schema validation. Typing has a lot of value, but don&amp;rsquo;t overdo it. It&amp;rsquo;s best used at fragile &lt;em&gt;boundaries&lt;/em&gt;. &lt;a href=&#34;https://chatgpt.com/share/69770709-c238-8003-899a-b8286a0e5474&#34;&gt;ChatGPT&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;⭐ Notes from &lt;a href=&#34;https://hollisrobbinsanecdotal.substack.com/p/llm-poetry-and-the-greatness-question&#34;&gt;LLM poetry and the &amp;ldquo;greatness&amp;rdquo; question&lt;/a&gt;:
&lt;ul&gt;
&lt;li&gt;Gwern follows this process to create good poetry. It&amp;rsquo;s a good structure for ANY kind of expert workflow with LLMs today:
&lt;ol&gt;
&lt;li&gt;&lt;strong&gt;Analyze&lt;/strong&gt; the style, content, and intent of the original.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Brainstorm 10+ different directions&lt;/strong&gt; the poem could go. Emphasize diversity.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Critique each&lt;/strong&gt; direction. Rate 1-5 stars.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Write the best&lt;/strong&gt; one.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Critique and edit line by line&lt;/strong&gt;.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Generate a new clean draft&lt;/strong&gt;.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Repeat&lt;/strong&gt; at least twice.&lt;/li&gt;
&lt;li&gt;Print final version.&lt;/li&gt;
&lt;/ol&gt;
&lt;/li&gt;
&lt;li&gt;&amp;ldquo;As a poet and scholar of poetry I feel comfortable arguing that Gwern’s work engineering prompts is, in effect, writing poetry.&amp;rdquo;&lt;/li&gt;
&lt;li&gt;Mercor uses expert poets to creates rubric. Models generate poem that experts grades, which refines the rubric, which trains the model.&lt;/li&gt;
&lt;li&gt;But models tend to the mean and need nudges (from humans?) to surface outliers and ascribe meaning (uniquely human?), which is where greatness lies.&lt;/li&gt;
&lt;/ul&gt;
&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://bsky.app/profile/emollick.bsky.social/post/3m5abr5vs5s23&#34;&gt;Ethan Mollick&lt;/a&gt;: &amp;ldquo;I keep warning that so many of our systems are still built around the assumption that quality writing and analysis are costly and therefore meaningful signals. Our systems are very much not ready for the revelation that this is no longer true, as this planning objection AI shows.&amp;rdquo;
Basically, AI lowers the cost of Government and Corporate interactions. It&amp;rsquo;d be a cool hack to agent-ify these to death, i.e. do all kinds of Government / Corporate interactions that were painful earlier, but now are much easier.&lt;/li&gt;
&lt;li&gt;I just realized: &amp;ldquo;Will AI take my job?&amp;rdquo; is a variant of &amp;ldquo;Will immigrants take my job?&amp;rdquo; or &amp;ldquo;Will affirmative action take my job?&amp;rdquo; &lt;em&gt;Any&lt;/em&gt; increase in labor capacity is a threat.
But then, the &lt;strong&gt;only&lt;/strong&gt; way to get promoted is if someone takes your job. So, maybe we should ask: &amp;ldquo;How do I become their boss?&amp;rdquo; Better yet, tell your boss &amp;ldquo;I created a 4-agent team and got 2X done. Give me a new title.&amp;rdquo;&lt;/li&gt;
&lt;li&gt;Some simple yet powerful AI adoption principles from &lt;a href=&#34;https://lethain.com/company-ai-adoption/&#34;&gt;Will Larson&lt;/a&gt; - that I&amp;rsquo;ve seen work rather well:
&lt;ul&gt;
&lt;li&gt;Make tools accessible&lt;/li&gt;
&lt;li&gt;Document tips &amp;amp; tricks&lt;/li&gt;
&lt;li&gt;Highlight how people (especially senior leaders) are using it&lt;/li&gt;
&lt;/ul&gt;
&lt;/li&gt;
&lt;li&gt;An &lt;a href=&#34;https://www.playbookatlas.com/research/ai-adoption-explorer&#34;&gt;analysis of 1,250 Claude user interviews&lt;/a&gt; indicates that:
&lt;ul&gt;
&lt;li&gt;Adoption of Creatives &amp;gt; Workforce &amp;gt; Scientists. Interestingly, the identity threat and guilt of Creatives &amp;gt; Workforce &amp;gt; Scientists!&lt;/li&gt;
&lt;li&gt;Creatives they feel they&amp;rsquo;re cheating, lazy, or not adding value! Scientists use it less, but it&amp;rsquo;s more a tool and THEY verify.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Sceptical verification&lt;/strong&gt; is the strongest thread.&lt;/li&gt;
&lt;/ul&gt;
&lt;/li&gt;
&lt;li&gt;Mintlify is &lt;a href=&#34;https://www.mintlify.com/blog/skill-md&#34;&gt;proposing &lt;code&gt;.well-known/skills/&lt;/code&gt;&lt;/a&gt; as the directory to store LLM skills sites want to publish. This could be an extension of the &lt;code&gt;llms.txt&lt;/code&gt; mechanism.&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://www.openresponses.org/&#34;&gt;Open Responses&lt;/a&gt; is the open version of OpenAI&amp;rsquo;s Responses API. &lt;a href=&#34;https://openrouter.ai/docs/api/reference/responses/overview&#34;&gt;OpenRouter&lt;/a&gt; and &lt;a href=&#34;https://huggingface.co/blog/open-responses&#34;&gt;HuggingFace&lt;/a&gt; support is a big deal, and though Google, Anthropic, Meta etc. don&amp;rsquo;t yet support it, they might.&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://rest.sh/&#34;&gt;Restish&lt;/a&gt; converts OpenAPI specs into CLI tools - with shell completion. Combined with an OAuth CLI like &lt;a href=&#34;https://github.com/SecureAuthCorp/oauth2c&#34;&gt;oauth2c&lt;/a&gt; this is a great way to conert APIs to CLI commands. &lt;a href=&#34;https://walters.app/blog/composing-apis-clis&#34;&gt;Via&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;Vercel&amp;rsquo;s &lt;a href=&#34;https://github.com/vercel-labs/agent-browser&#34;&gt;agent-browser&lt;/a&gt; seems a good CLI choice for browser automation, alongside &lt;a href=&#34;https://github.com/microsoft/playwrite-cli&#34;&gt;playwright-cli&lt;/a&gt;. It may be work switching from direct Playwright coding (on CDP). &lt;a href=&#34;https://chatgpt.com/share/69770999-5d50-8003-8795-d297a1bb0c09&#34;&gt;ChatGPT&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;Capturing actions using HAR and passing it to LLMs seems like another clever way of using AI coding agents for browser automation. &lt;a href=&#34;https://walters.app/blog/composing-apis-clis&#34;&gt;Via&lt;/a&gt;
&lt;ul&gt;
&lt;li&gt;Open a browser.&lt;/li&gt;
&lt;li&gt;Open Devtools &amp;gt; Network and filter to HTML, XHR, WS, Other.&lt;/li&gt;
&lt;li&gt;Do what you want to automate, i.e. load LinkedIn, search, scroll, fetch next pages, etc.&lt;/li&gt;
&lt;li&gt;Devtools &amp;gt; Network &amp;gt; right click &amp;gt; “Save All As HAR”.&lt;/li&gt;
&lt;li&gt;Run the file through a HAR-sanitizer&lt;/li&gt;
&lt;li&gt;Prompt: “Create a Python client to automate the actions I captured in file.har&amp;quot;.&lt;/li&gt;
&lt;/ul&gt;
&lt;/li&gt;
&lt;li&gt;When any AI coding agent can build apps, value will probably migrate away from software to data, network (distribution and users), trust, taste, and physical goods. Owning these controls value. Also, infrastructure to run vibe-coded apps (e.g. auth, hosting, DB, LLM APIs, etc. bundled) will likely lead to Medium / WordPress like platforms.&lt;/li&gt;
&lt;li&gt;After 30 years of learning (and teaching) statistics, I finally found a good explanation of R². R²=80% means that ~80% of the change is because of the other variable. &lt;a href=&#34;https://gemini.google.com/share/f3dfa6cfaf89&#34;&gt;Gemini&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;⭐ People think numbers create trust; often they create attack surfaces.
&lt;ul&gt;
&lt;li&gt;Goodhart&amp;rsquo;s Law: &amp;ldquo;When a measure becomes a target, it ceases to be a good measure.&amp;rdquo; By providing a number, you invite people to &amp;ldquo;game&amp;rdquo; the system or find the flaws in how that number was manufactured.&lt;/li&gt;
&lt;li&gt;The Precision Trap: While precise numbers can increase &lt;em&gt;perceived&lt;/em&gt; credibility initially, they also lead to &amp;ldquo;anchoring.&amp;rdquo; If the number is even slightly off, the entire foundation of trust collapses more violently than it would for a general estimate.&lt;/li&gt;
&lt;li&gt;Statistical Literacy Gap: Most people don&amp;rsquo;t argue with &amp;ldquo;vibes,&amp;rdquo; but many will argue with &amp;ldquo;averages&amp;rdquo; if their personal experience represents an outlier. The number creates a surface for anecdotal rebuttal.&lt;/li&gt;
&lt;/ul&gt;
&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://www.eraser.io/&#34;&gt;Eraser.io&lt;/a&gt; offers an &lt;a href=&#34;https://www.eraser.io/ai/architecture-diagram-generator&#34;&gt;AI architecture diagram generator&lt;/a&gt; that creates reasonable architectures. It uses its own &lt;a href=&#34;https://www.eraser.io/guides/best-diagram-as-code-tools-in-2025&#34;&gt;diagram-as-code DSL&lt;/a&gt;, competing with &lt;a href=&#34;https://d2lang.com/&#34;&gt;D2&lt;/a&gt;, &lt;a href=&#34;https://plantuml.com/&#34;&gt;PlantUML&lt;/a&gt;, &lt;a href=&#34;https://mermaid-js.github.io/&#34;&gt;Mermaid&lt;/a&gt;,&lt;/li&gt;
&lt;li&gt;Exposing your workflow as a software interface productizes services businesses. For example, my auditors and immigration lawyers have portals where I can fill out forms, upload documents, see my status, etc. This standardizes their delivery, and creates a &amp;ldquo;product&amp;rdquo; moat.&lt;/li&gt;
&lt;li&gt;⭐ Your &amp;ldquo;villains&amp;rdquo; or enemies are often alternatives/backups that have a role in the ecosystem, offering diversity/resilience when you&amp;rsquo;re wrong. Create roles and incentives for them rather than eliminating them. For example:
&lt;ul&gt;
&lt;li&gt;Don&amp;rsquo;t make LLMs do all the work. Create a role for the clunky SQL whose resilience saves the day when LLMs hallucinate.&lt;/li&gt;
&lt;li&gt;Make the person who hates your prototype the Red Team Lead - to catch the flaws you miss.&lt;/li&gt;
&lt;li&gt;Make the people who reject your product the scouts / innovators - to find alternatives you miss.&lt;/li&gt;
&lt;/ul&gt;
&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://neon.com&#34;&gt;Neon.com&lt;/a&gt; is like Supabase but without auth, functions, etc. It&amp;rsquo;s just Postgres as a service. An alternative for prototypes (that I haven&amp;rsquo;t tried yet.) &lt;a href=&#34;https://chatgpt.com/share/69763727-0628-8003-9a99-896a055a8b6e&#34;&gt;ChatGPT&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://supertokens.com/&#34;&gt;SuperTokens&lt;/a&gt; is an open-source self-hosted auth service that I&amp;rsquo;m hearing about more often, but haven&amp;rsquo;t tested. Seems to be ahead of alternatives like &lt;a href=&#34;https://authjs.dev/&#34;&gt;Auth.js&lt;/a&gt; / &lt;a href=&#34;https://www.better-auth.com/&#34;&gt;Better Auth&lt;/a&gt;. &lt;a href=&#34;https://chatgpt.com/share/69763779-9150-8003-bb9f-081685c99dc2&#34;&gt;ChatGPT&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://kontinentalist.com/stories/bollywood-falls-out-of-love&#34;&gt;Bollywood Falls Out Of Love&lt;/a&gt; is a great visual data story on &lt;a href=&#34;https://kontinentalist.com/&#34;&gt;The Kontinentalist&lt;/a&gt; by &lt;a href=&#34;https://www.linkedin.com/in/surbhi-bhatia/&#34;&gt;Surbhi&lt;/a&gt; about the decline of romance and growth of nationalism on bollywood genres.&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://recharts.github.io/&#34;&gt;Recharts&lt;/a&gt; is a React charting library with some &lt;a href=&#34;https://recharts.github.io/en-US/examples/&#34;&gt;slick capabilities&lt;/a&gt; like &lt;a href=&#34;https://recharts.github.io/en-US/storybook/&#34;&gt;brushing&lt;/a&gt;, &lt;a href=&#34;https://recharts.github.io/en-US/storybook/&#34;&gt;customizable tooltips&lt;/a&gt;, and &lt;a href=&#34;https://www.dataforindia.com/population-growth/&#34;&gt;bar chart races&lt;/a&gt;. Via &lt;a href=&#34;https://www.linkedin.com/in/rukmini-s-553483b9/&#34;&gt;Rukmini&lt;/a&gt; - &lt;a href=&#34;https://www.dataforindia.com/&#34;&gt;Data for India&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://github.com/QwenLM/Qwen3-TTS&#34;&gt;Qwen3 TTS&lt;/a&gt; is impressive. It voice-clones, streams, and the tone/style can be controlled via prompts. The model is small. I ran it locally without &lt;code&gt;flash-attn&lt;/code&gt; (which I couldn&amp;rsquo;t get to work) and took ~14 seconds to generate an audio file for 10 words on my GPU machine. Environment setup:
&lt;div class=&#34;highlight&#34;&gt;&lt;pre tabindex=&#34;0&#34; class=&#34;chroma&#34;&gt;&lt;code class=&#34;language-bash&#34; data-lang=&#34;bash&#34;&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;uv venv --python 3.12
&lt;/span&gt;&lt;/span&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;&lt;span class=&#34;nv&#34;&gt;UV_TORCH_BACKEND&lt;/span&gt;&lt;span class=&#34;o&#34;&gt;=&lt;/span&gt;auto uv pip install -U qwen-tts
&lt;/span&gt;&lt;/span&gt;&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;/li&gt;
&lt;li&gt;DeepSeek created an external memory system for LLMs that lets them look up (instead of computing to remember) knowledge. That means CPU RAM can be used instead of GPU, models can become smaller, and training can become faster. This looks like an example of how algorithms/ideas can continue the scaling laws. &lt;a href=&#34;https://gemini.google.com/share/a94760cc5e2e&#34;&gt;Gemini&lt;/a&gt; via &lt;a href=&#34;https://bsky.app/profile/eugenevinitsky.bsky.social/post/3mcap4nt5ms2g&#34;&gt;Jeremy Howard&lt;/a&gt;&lt;/li&gt;
&lt;/ul&gt;
</description>
    </item>
    <item>
      <title>Things I Learned - 31 Dec 2023</title>
      <link>https://www.s-anand.net/blog/things-i-learned-31-dec-2023/</link>
      <pubDate>Sun, 31 Dec 2023 00:00:00 +0000</pubDate>
      <guid>https://www.s-anand.net/blog/things-i-learned-31-dec-2023/</guid>
      <description>&lt;p&gt;This week, I learned:&lt;/p&gt;
&lt;ul&gt;
&lt;li&gt;Quantum computing is slow, has low transfer bandwidths, and only prime factorization has an exponentially faster algorithm. &lt;a href=&#34;https://spectrum.ieee.org/quantum-computing-skeptics&#34;&gt;via&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;The hidden brain podcast. What would Socrates do? Also Philosophy Bites Podcast: why do philosophers use example. And: the happiness lab: happiness lessons of the ancients
&lt;ul&gt;
&lt;li&gt;How many of our beliefs are truly our own? How many are a product of our environment? Contrast these and identify your true beliefs&lt;/li&gt;
&lt;li&gt;For every thought and action you have, even tiny ones, ask &amp;ldquo;Why am I doing that?&amp;rdquo; Dig deeper because it may not be intrinsic&lt;/li&gt;
&lt;li&gt;One way to become memorable is to.write stuff others will reproduce for a long time. Plato and Aristotle did that&lt;/li&gt;
&lt;li&gt;everyone has multiple personality. This is partly because different parts of the brain evolved independently for different functions. System one and system two thinking are just such one broad classification. e.g. We think our train is moving when the nearby train moves because our visual brain is faster than our somatic brain.&lt;/li&gt;
&lt;li&gt;Good lessons and pitches cater to the rational AND the subconscious. Reason AND story. To activate different parts of the brain. That&amp;rsquo;s why philosophers use examples&lt;/li&gt;
&lt;li&gt;Philosophy brings change through reason. Revelations: through sudden insight. Rhetoric: through insight.&lt;/li&gt;
&lt;li&gt;Act as if you already are what you want to become. Aristotle&lt;/li&gt;
&lt;li&gt;Align your environment (including habits) to your beliefs. It will become easier to act your beliefs then.&lt;/li&gt;
&lt;li&gt;All virtues are moderation. It&amp;rsquo;s possible to take every virtue to the wrong extreme&lt;/li&gt;
&lt;li&gt;Some Christians have wristband that reads WWJD. What would Jesus do? Explore yourself a reminder of what would X do. Maybe Benjamin Franklin, Socrates, Feynman, etc&lt;/li&gt;
&lt;li&gt;People mistake their environment for their feelings. 1970s Experiment: People on a shaky bridge think they love each other. Experiment: people rationalize things irrespective of reality.&lt;/li&gt;
&lt;li&gt;&amp;ldquo;The Unexamined Life&amp;rdquo; is about questioning theories or stories or maps constantly. It&amp;rsquo;s also about questioning our thoughts and emotions constantly. Mindfulness is the VERBAL way of doing this. Meditation is the NON-VERBAL way of paying attention. Both are Processes to remove distraction and increase authenticity.&lt;/li&gt;
&lt;li&gt;Learning about people is a good way to learn about ourselves. And vice versa.&lt;/li&gt;
&lt;/ul&gt;
&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://www.lica.world/&#34;&gt;Lica&lt;/a&gt; has a fascinating demo of how a document can be converted into a video story.&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://www.youtube.com/watch?v=qEDvmvBbDBk&#34;&gt;Spillnot&lt;/a&gt; doesn&amp;rsquo;t spill drinks even when you swing!&lt;/li&gt;
&lt;li&gt;Things super-intelligences could do that humans can&amp;rsquo;t:
&lt;ul&gt;
&lt;li&gt;Solving complex mathematical problems&lt;/li&gt;
&lt;li&gt;Advanced scientific discovery (quantum computing, nanotechnology, biotechnology)&lt;/li&gt;
&lt;li&gt;Ultra-precise predictive modeling in complex systems (climate, economics, social dynamics)&lt;/li&gt;
&lt;li&gt;Optimizing global systems at high precision (logistics, traffic, energy distribution, resource allocation)&lt;/li&gt;
&lt;li&gt;Universal translation (unknown languages, animal communication, extraterrestrial signals)&lt;/li&gt;
&lt;li&gt;Deep medical personalization: individualized medical treatments from genetics, environment, and lifestyle&lt;/li&gt;
&lt;li&gt;Create new materials: Designing materials or chemicals with specific properties&lt;/li&gt;
&lt;li&gt;Complex system integration: combining AI, bio tech, nano tech in new ways&lt;/li&gt;
&lt;li&gt;Philosophical insights: new perspectives or solutions to age-old philosophical dilemmas&lt;/li&gt;
&lt;li&gt;Space exploration and colonization&lt;/li&gt;
&lt;li&gt;Predicting natural disasters&lt;/li&gt;
&lt;li&gt;Customized education at scale&lt;/li&gt;
&lt;/ul&gt;
&lt;/li&gt;
&lt;li&gt;Ways of working with them
&lt;ul&gt;
&lt;li&gt;Collaborative problem solving&lt;/li&gt;
&lt;li&gt;Creative collaboration&lt;/li&gt;
&lt;li&gt;Decision support&lt;/li&gt;
&lt;li&gt;Personalized education&lt;/li&gt;
&lt;li&gt;Establishing ethical and safety protocols&lt;/li&gt;
&lt;li&gt;Recreational and leisure activities&lt;/li&gt;
&lt;/ul&gt;
&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://arxiv.org/abs/2312.12682&#34;&gt;Mini-GPTs&lt;/a&gt; is an interesting approach to shrink LLMs and make them domain specific. It takes existing LLMs and removes neurons not used in a specific domain (e.g. law, medicine, etc.)&lt;/li&gt;
&lt;li&gt;Book to read (again) about how to take a team beyond their abilities even if you&amp;rsquo;re not the expert
&lt;ul&gt;
&lt;li&gt;&amp;ldquo;Measure What Matters&amp;rdquo; by John Doerr&lt;/li&gt;
&lt;li&gt;&amp;ldquo;High Output Management&amp;rdquo; by Andy Grove&lt;/li&gt;
&lt;li&gt;&amp;ldquo;The Checklist Manifesto&amp;rdquo; by Atul Gawande&lt;/li&gt;
&lt;li&gt;&amp;ldquo;The Lean Startup&amp;rdquo; by Eric Ries&lt;/li&gt;
&lt;li&gt;&amp;ldquo;Creativity, Inc&amp;rdquo; by Ed Catmull&lt;/li&gt;
&lt;li&gt;&amp;ldquo;The Hard Thing About Hard Things&amp;rdquo; by Ben Horowitz&lt;/li&gt;
&lt;li&gt;&amp;ldquo;The Four Disciplines of Execution&amp;rdquo; by Chris McChesney, Sean Covey, and Jim Huling&lt;/li&gt;
&lt;/ul&gt;
&lt;/li&gt;
&lt;/ul&gt;
</description>
    </item>
    <item>
      <title>Socratic method</title>
      <link>https://www.s-anand.net/blog/socratic-method/</link>
      <pubDate>Fri, 26 Jan 2001 12:00:00 +0000</pubDate>
      <guid>https://www.s-anand.net/blog/socratic-method/</guid>
      <description>&lt;p&gt;Rick Garlikov tried using the &lt;a href=&#34;http://www.garlikov.com/Soc_Meth.html&#34;&gt;Socratic method&lt;/a&gt; to teach binary numbers to a third grade class. Looks like it worked well. I&amp;rsquo;m all for the Socratic method of teaching.&lt;/p&gt;
</description>
    </item>
  </channel>
</rss>
