<?xml version="1.0" encoding="utf-8" standalone="yes"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:content="http://purl.org/rss/1.0/modules/content/">
  <channel>
    <title>ai-workflows on S Anand</title>
    <link>https://www.s-anand.net/blog/tag/ai-workflows/</link>
    <description>Recent content in ai-workflows on S Anand</description>
    <generator>Hugo -- 0.164.0</generator>
    <language>en-us</language>
    <lastBuildDate>Tue, 04 Aug 2026 05:48:31 +0800</lastBuildDate>
    <atom:link href="https://www.s-anand.net/blog/tag/ai-workflows/index.xml" rel="self" type="application/rss+xml" />
    <item>
      <title>AI tax returns 2026</title>
      <link>https://www.s-anand.net/blog/ai-tax-returns-2026/</link>
      <pubDate>Tue, 04 Aug 2026 05:48:31 +0800</pubDate>
      <guid>https://www.s-anand.net/blog/ai-tax-returns-2026/</guid>
      <description>&lt;p&gt;&lt;img loading=&#34;lazy&#34; src=&#34;https://files.s-anand.net/images/2026-06-05-ai-tax-returns-2026.avif&#34;&gt;&lt;/p&gt;
&lt;p&gt;On 16 July, my auditor sent me a draft Indian tax return: a refund of Rs 3 lakhs.&lt;/p&gt;
&lt;p&gt;I gave ChatGPT my AIS, Form 26AS, bank statements, mutual fund statements, property papers, travel records, prior returns, and so on, told it &lt;em&gt;not&lt;/em&gt; to look at the auditor&amp;rsquo;s draft, and asked it to calculate my tax independently.&lt;/p&gt;
&lt;p&gt;It calculated a refund of about Rs 2.8 lakhs, roughly Rs 20K lower. (Less money, but a smaller refund felt less worse than a larger tax.)&lt;/p&gt;
&lt;p&gt;Why? The auditor treated a Rs 40K in Form 26AS as &lt;em&gt;interest&lt;/em&gt; on an income-tax refund. ChatGPT said (based on an earlier tax intimation) that it was &lt;em&gt;tax&lt;/em&gt; deducted on an interest of Rs 1.3 lakhs.&lt;/p&gt;
&lt;p&gt;My auditor disagreed, so we got on a screen-sharing call. She showed me Form 26AS. I showed her the earlier tax intimation. (It &lt;em&gt;was&lt;/em&gt; confusing.)&lt;/p&gt;
&lt;p&gt;&amp;ldquo;I think it was a mistake. I&amp;rsquo;ll rectify it,&amp;rdquo; she said.&lt;/p&gt;
&lt;hr&gt;
&lt;p&gt;I had a large capital gain from redeeming Indian mutual funds. My auditor said it&amp;rsquo;s taxable in India.&lt;/p&gt;
&lt;p&gt;ChatGPT found a similar case, &lt;a href=&#34;https://indiankanoon.org/doc/23391260/&#34;&gt;&lt;em&gt;Anushka Sanjay Shah v. ITO&lt;/em&gt;&lt;/a&gt;. A Singapore tax resident had redeemed Indian mutual funds. The court said that the gain was taxable only in Singapore under the India-Singapore tax treaty.&lt;/p&gt;
&lt;p&gt;My auditor said &amp;ldquo;No, only NRE accounts are exempt, not NRO accounts&amp;rdquo;. ChatGPT said &amp;ldquo;But the judgement does not mention the type of account at all&amp;rdquo;, which I told my auditor.&lt;/p&gt;
&lt;p&gt;She checked with a senior colleague. The next morning she called. &amp;ldquo;You are eligible&amp;rdquo;, she said, &amp;ldquo;I will file it based on that scenario.&amp;rdquo; (I asked for that in writing.)&lt;/p&gt;
&lt;hr&gt;
&lt;p&gt;Yes, AI got me a better (or more correct) return. But frankly, it could have gone either way (been wrong, gotten me a worse return).&lt;/p&gt;
&lt;p&gt;The main lesson I&amp;rsquo;m taking away is: &lt;strong&gt;The auditor cross-checked AI and filed it as their responsibility.&lt;/strong&gt;&lt;/p&gt;
&lt;p&gt;The useful Expert + AI workflow, for me, is:&lt;/p&gt;
&lt;ol&gt;
&lt;li&gt;Let AI calculate independently without the expert.&lt;/li&gt;
&lt;li&gt;Give it raw sources, not summaries. (It can crunch lots more data.)&lt;/li&gt;
&lt;li&gt;Ask the expert precise questions with evidence.&lt;/li&gt;
&lt;li&gt;Get the expert&amp;rsquo;s position in writing for important things.&lt;/li&gt;
&lt;li&gt;Let the expert take the final call.&lt;/li&gt;
&lt;/ol&gt;
&lt;p&gt;That division worked rather well, I think.&lt;/p&gt;
&lt;!--
Source chats:
- https://chatgpt.com/c/6a5eb53c-a954-83e8-bab1-effbcb70ced2 Anand Tax AY2026-27 Calculation
- https://chatgpt.com/c/6a6ae0bf-ff18-83ec-83dd-962a126bb6b8 Anand Tax AY2026-27 Calculation v2
- https://chatgpt.com/c/6a706237-a2c4-83ec-aa51-ea6fed960b50 Tax Filing Sequence Analysis
--&gt;
</description>
    </item>
    <item>
      <title>IIM Alumni AI Workflows Workshop</title>
      <link>https://www.s-anand.net/blog/iim-alumni-ai-workflows-workshop/</link>
      <pubDate>Sun, 21 Jun 2026 13:01:47 +0800</pubDate>
      <guid>https://www.s-anand.net/blog/iim-alumni-ai-workflows-workshop/</guid>
      <description>&lt;p&gt;The theme of yesterday&amp;rsquo;s workshop for the IIM Alumni at Singapore was &lt;strong&gt;Tools and Workflows&lt;/strong&gt; was:&lt;/p&gt;
&lt;p&gt;Agents are getting smarter, so they know what to do.&lt;br&gt;
Tools agents can use are growing and are more powerful.&lt;br&gt;
This combinatorial explosion creates explosive possibilites.&lt;/p&gt;
&lt;p&gt;This workshop covered the following six workflows:&lt;/p&gt;
&lt;ol&gt;
&lt;li&gt;&lt;strong&gt;Leverage transcripts.&lt;/strong&gt; Use &lt;a href=&#34;https://aistudio.google.com/prompts/new%5Fchat&#34;&gt;Google AI Studio&lt;/a&gt; to transcribe non-sensitive recordings with a reusable &lt;a href=&#34;https://github.com/sanand0/blog/blob/9763887917a006c8aa76b4fc60ed2202548bab6f/pages/prompts/transcribe-talk.md&#34;&gt;&amp;ldquo;don&amp;rsquo;t miss anything&amp;rdquo;&lt;/a&gt; prompt. AI Studio&amp;rsquo;s record button is a ready-to-use transcriber.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Simplify dense text&lt;/strong&gt; as a &lt;a href=&#34;https://www.s-anand.net/blog/creating-comic-explainers/&#34;&gt;comic&lt;/a&gt;, an &lt;a href=&#34;https://notebooklm.google.com/&#34;&gt;infographic&lt;/a&gt;, a &lt;a href=&#34;https://talks.s-anand.net/2026-05-23-ai-unboxed-context-engineering/&#34;&gt;story&lt;/a&gt;. Image generation is now a tool call an agent runs for you. Then &lt;a href=&#34;https://squoosh.app/&#34;&gt;compress it as AVIF on Squoosh&lt;/a&gt; before you email it to a thousand people.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Verify - cheaply.&lt;/strong&gt; Paste one suffix: &lt;em&gt;&amp;ldquo;Break this into key claims, mark certainty, flag the five highest-risk ones, and tell me how to verify or falsify each.&amp;rdquo;&lt;/em&gt; Convert to a &lt;a href=&#34;https://skill.md/&#34;&gt;skill&lt;/a&gt; to automate. Cross-checking with multiple models took &lt;a href=&#34;https://sanand0.github.io/llmevals/double-checking/&#34;&gt;error from 14% to 0.7%&lt;/a&gt;.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Skills are assets.&lt;/strong&gt; A &lt;a href=&#34;https://www.anthropic.com/engineering/equipping-agents-for-the-real-world-with-agent-skills&#34;&gt;skill&lt;/a&gt; tells the agent &amp;ldquo;here&amp;rsquo;s how &lt;em&gt;I&lt;/em&gt; do stuff.&amp;rdquo; Build them slowly, edit them weekly, and they compound for years. No skills support in your tool? Keep them as copy-pasteable prompts.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Brainstorm by forcing range.&lt;/strong&gt; &lt;a href=&#34;https://github.com/sanand0/scripts/blob/d9fde4f5f169c1d465603e0cd6095db3fcdabad3/agents/ideation-protocol/SKILL.md&#34;&gt;Ban the five obvious ideas; borrow from unrelated domains&lt;/a&gt;; smash two random concepts together with the &lt;a href=&#34;https://tools.s-anand.net/ideator/&#34;&gt;Ideator&lt;/a&gt;. Hallucination is a feature when you&amp;rsquo;re being creative.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;&lt;a href=&#34;https://help.openai.com/en/articles/10291617-scheduled-tasks-in-chatgpt&#34;&gt;Schedule tasks&lt;/a&gt;.&lt;/strong&gt; Weekly regulatory scans, daily meeting prep, market briefings - and even an &lt;a href=&#34;https://www.s-anand.net/blog/prompts/unreasonable-gesture/&#34;&gt;&amp;ldquo;unreasonable gesture&amp;rdquo;&lt;/a&gt; nudge. As AI hides the tech, human relationships gain value.&lt;/li&gt;
&lt;/ol&gt;
&lt;p&gt;Here&amp;rsquo;s the &lt;a href=&#34;https://youtu.be/H3GmGl1hzxM&#34;&gt;talk video&lt;/a&gt; and &lt;a href=&#34;https://talks.s-anand.net/2026-06-20-ai-unboxed-tools-workflows/&#34;&gt;full story + transcript&lt;/a&gt;.&lt;/p&gt;
&lt;p&gt;&lt;img loading=&#34;lazy&#34; src=&#34;https://talks.s-anand.net/2026-06-20-ai-unboxed-tools-workflows/comic-page.avif&#34;&gt;&lt;/p&gt;
</description>
    </item>
    <item>
      <title>Repurposing blog posts for talks</title>
      <link>https://www.s-anand.net/blog/repurposing-blog-posts-for-talks/</link>
      <pubDate>Sun, 22 Feb 2026 22:09:49 +0800</pubDate>
      <guid>https://www.s-anand.net/blog/repurposing-blog-posts-for-talks/</guid>
      <description>&lt;p&gt;Recently, I&amp;rsquo;ve re-used my own writing / transcripts as context to LLMs. For example, I&amp;rsquo;ve used:&lt;/p&gt;
&lt;ul&gt;
&lt;li&gt;&lt;a href=&#34;https://www.s-anand.net/blog/transcript-ai-ded-interviews/&#34;&gt;My meeting transcripts to answer interview questions&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://www.s-anand.net/blog/writing-articles-from-my-blog-posts/&#34;&gt;My blog posts to write news articles&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://www.s-anand.net/blog/extracting-ai-advice/&#34;&gt;My chat history to extract AI-related advice&lt;/a&gt;&lt;/li&gt;
&lt;/ul&gt;
&lt;p&gt;This repurposing can be used for so many things.&lt;/p&gt;
&lt;p&gt;For example, before delivering a talk to journalists &lt;em&gt;&amp;ldquo;Review my Feb 2026 LLM posts and generate a single-sentence, ELI15 high-impact use case for journalists.&amp;rdquo;&lt;/em&gt; gets me list of use cases. Now, all I have to do is &lt;strong&gt;show what I did&lt;/strong&gt; and &lt;strong&gt;share how it&amp;rsquo;s relevant&lt;/strong&gt; for them, like:&lt;/p&gt;
&lt;ol&gt;
&lt;li&gt;&lt;a href=&#34;https://www.s-anand.net/blog/finding-old-friends-with-gemini/&#34;&gt;I found old friends with Gemini Deep Research&lt;/a&gt;. You can trace sources who changed names.&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://www.s-anand.net/blog/gemini-3-flash-ocrs-dilbert-accurately/&#34;&gt;I transcribed the entire Dilbert archive&lt;/a&gt;. You can transcribe scanned court records for $20.&lt;/li&gt;
&lt;li&gt;&amp;hellip; and so on.&lt;/li&gt;
&lt;/ol&gt;
&lt;p&gt;I can do this for &lt;em&gt;any field&lt;/em&gt;.&lt;/p&gt;
&lt;p&gt;The approach is simple. I copy recent blog posts related to LLMs via:&lt;/p&gt;
&lt;div class=&#34;highlight&#34;&gt;&lt;pre tabindex=&#34;0&#34; class=&#34;chroma&#34;&gt;&lt;code class=&#34;language-bash&#34; data-lang=&#34;bash&#34;&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;rg -l &lt;span class=&#34;s2&#34;&gt;&amp;#34;^[[:space:]]*- llms&amp;#34;&lt;/span&gt; -g &lt;span class=&#34;s2&#34;&gt;&amp;#34;*.md&amp;#34;&lt;/span&gt; &lt;span class=&#34;p&#34;&gt;|&lt;/span&gt; xargs rg &lt;span class=&#34;s2&#34;&gt;&amp;#34;^date:&amp;#34;&lt;/span&gt; &lt;span class=&#34;p&#34;&gt;|&lt;/span&gt; sort -k2 -r &lt;span class=&#34;p&#34;&gt;|&lt;/span&gt; head -n &lt;span class=&#34;m&#34;&gt;30&lt;/span&gt; &lt;span class=&#34;p&#34;&gt;|&lt;/span&gt; cut -d: -f1 &lt;span class=&#34;p&#34;&gt;|&lt;/span&gt; uniq &lt;span class=&#34;p&#34;&gt;|&lt;/span&gt; xargs uvx files-to-prompt --cxml &lt;span class=&#34;p&#34;&gt;|&lt;/span&gt; xclip -selection clipboard
&lt;/span&gt;&lt;/span&gt;&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;p&gt;&amp;hellip; and ask the LLM:&lt;/p&gt;
&lt;blockquote&gt;
&lt;p&gt;Pick articles with the best application to [FIELD]. For each, write a single-sentence, ELI15 use case DIRECTLY following from the article. Focus on high-impact, common scenarios solving a specific problem (a persona trying to do X but blocked by Y). Skip articles without a clear, practical application. Be ruthlessly concise: ~40 words per use case. Format each line as: &lt;code&gt;Article Title: Use case description&lt;/code&gt;.&lt;/p&gt;
&lt;/blockquote&gt;
&lt;p&gt;&lt;img loading=&#34;lazy&#34; src=&#34;https://files.s-anand.net/images/2026-02-22-repurposing-blog-posts-for-talks.avif&#34;&gt; &lt;!-- https://gemini.google.com/u/2/app/407bc3134d719117 --&gt;&lt;/p&gt;
&lt;p&gt;Here are a few examples (and they&amp;rsquo;re &lt;em&gt;good&lt;/em&gt; ones):&lt;/p&gt;
&lt;details&gt;
&lt;summary&gt;See journalist use cases&lt;/summary&gt;
&lt;ol&gt;
&lt;li&gt;&lt;a href=&#34;https://www.s-anand.net/blog/extracting-ai-advice/&#34;&gt;Extracting AI Advice&lt;/a&gt;: A journalist who has obtained hundreds of FOI documents or deposition transcripts can use a cheap large-context model to pull one-sentence bullets from each and then ask a stronger model to rank the top patterns &amp;ndash; turning months of reading into an afternoon.&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://www.s-anand.net/blog/finding-old-friends-with-gemini/&#34;&gt;Finding Old Friends with Gemini&lt;/a&gt;: An investigative journalist can use Gemini Deep Research to trace sources who have changed names, employers, or countries &amp;ndash; surfacing government nomination lists and public records that bridge who someone used to be with who they are now.&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://www.s-anand.net/blog/gemini-3-flash-ocrs-dilbert-accurately/&#34;&gt;Gemini 3 Flash OCRs Dilbert Accurately&lt;/a&gt;: A journalist who receives a dump of scanned physical documents &amp;ndash; leaked government files, court records &amp;ndash; can make the entire archive full-text searchable for roughly $20, without waiting weeks for manual transcription.&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://www.s-anand.net/blog/organizing-pdf-receipts/&#34;&gt;Organizing PDF Receipts&lt;/a&gt;: An investigative journalist who obtains a dump of expense or procurement PDFs can ask an AI coding tool to parse each one and rename it to a standard format &amp;ndash; making it trivial to spot duplicate amounts, unusual vendors, or suspicious date gaps at a glance.&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://www.s-anand.net/blog/rip-data-engineers/&#34;&gt;RIP, Data Engineers&lt;/a&gt;: A data journalist who obtains the SQL query logs of a public institution can use an AI agent to cluster what questions the database was actually built to answer &amp;ndash; revealing whether an agency is doing tactical monitoring or genuine public-service analysis, which can itself be the story.&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://www.s-anand.net/blog/transcript-ai-ded-interviews/&#34;&gt;Transcript AI-ded Interviews&lt;/a&gt;: A journalist who has done twenty interviews for a long-form profile can feed all the transcripts to a large-context model and ask it to surface the best quotes on a given angle, or flag contradictions between what different sources said &amp;ndash; without re-reading every word.&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://www.s-anand.net/blog/using-ai-for-work-news/&#34;&gt;Using AI for Work News&lt;/a&gt;: A beat journalist without a research desk can set up a Google Workspace automation that scans their sources weekly and delivers a single Gemini-written brief &amp;ndash; so nothing from a regulator&amp;rsquo;s filing or a council report slips through unread.&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://www.s-anand.net/blog/writing-articles-from-my-blog-posts/&#34;&gt;Writing Articles from My Blog Posts&lt;/a&gt;: A reporter with years of stories on the same beat can feed their archive to an AI, ask it to identify which past pieces form the strongest foundation for a new investigation, and get a draft synthesis written in their own voice &amp;ndash; rather than starting from a blank page.&lt;/li&gt;
&lt;/ol&gt;
&lt;/details&gt;
&lt;details&gt;
&lt;summary&gt;See civic use cases&lt;/summary&gt;
&lt;ul&gt;
&lt;li&gt;&lt;a href=&#34;https://www.s-anand.net/blog/breaking-rules-in-the-age-of-ai/&#34;&gt;Breaking Rules in the Age of AI&lt;/a&gt;: A government adult literacy programme can drop its AI ban and replace it with instant AI-graded feedback &amp;ndash; letting learners ask questions in their own language, fail and retry freely, and delegate tedious steps &amp;ndash; re-engaging adults who had already given up on education.&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://www.s-anand.net/blog/extracting-ai-advice/&#34;&gt;Extracting AI Advice&lt;/a&gt;: A city welfare department can transcribe its 300 community consultations into one-sentence bullets and rank the top 10 concerns across meetings in the residents&amp;rsquo; words &amp;ndash; for a few dollars and hours.&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://www.s-anand.net/blog/finding-old-friends-with-gemini/&#34;&gt;Finding Old Friends with Gemini&lt;/a&gt;: A social welfare officer can use Gemini Deep Research to trace former programme participants who have moved or changed names, surfacing public records that link old and new identities &amp;ndash; closing the loop on services that were started but never finished.&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://www.s-anand.net/blog/gemini-3-flash-ocrs-dilbert-accurately/&#34;&gt;Gemini 3 Flash OCRs Dilbert Accurately&lt;/a&gt;: A city archivist can make thousands of scanned land records, court orders, and petitions fully searchable by running them through Gemini Flash at roughly $20 for an entire archive &amp;ndash; no in-person visit required.&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://www.s-anand.net/blog/organizing-pdf-receipts/&#34;&gt;Organizing PDF Receipts&lt;/a&gt;: An NGO finance coordinator can ask an AI coding tool to write a script that reads every vendor PDF, extracts the date, amount, and reference, and renames the file to a standard format &amp;ndash; turning a half-day filing chore into a five-minute step before uploading to a compliance portal.&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://www.s-anand.net/blog/rip-data-engineers/&#34;&gt;RIP, Data Engineers&lt;/a&gt;: A government department can expose its data culture &amp;ndash; and fix silent metric misalignment between teams &amp;ndash; by feeding its SQL query logs to an AI agent that clusters them and proposes a small set of shared standard tables, no data warehouse project required.&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://www.s-anand.net/blog/tds-comic-generation/&#34;&gt;TDS Comic Generation&lt;/a&gt;: A district health office with no design budget can generate multilingual public health comics for low-literacy communities by defining a few recurring characters, writing their dialogue in plain language, and letting Gemini produce each strip in minutes &amp;ndash; at near-zero cost per language.&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://www.s-anand.net/blog/transcript-ai-ded-interviews/&#34;&gt;Transcript AI-ded Interviews&lt;/a&gt;: A government communications officer can feed all auto-generated meeting transcripts on a topic to a large-context model and get a press-ready 150-word statement grounded in what the department actually said &amp;ndash; in an hour instead of a week of document hunting.&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://www.s-anand.net/blog/using-ai-for-work-news/&#34;&gt;Using AI for Work News&lt;/a&gt;: A district administrator receiving eight separate departmental reports can set up a 20-minute Google Workspace automation that delivers one weekly email surfacing cross-department clashes &amp;ndash; like a public works delay that will knock out a scheduled health camp &amp;ndash; that no individual report would flag.&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://www.s-anand.net/blog/writing-articles-from-my-blog-posts/&#34;&gt;Writing Articles from My Blog Posts&lt;/a&gt;: A civic think tank analyst with 40 research reports can ask AI to pick the strongest op-ed angle for a target publication and draft it from their own words &amp;ndash; in an afternoon, not a week of rewriting from scratch.&lt;/li&gt;
&lt;/ul&gt;
&lt;/details&gt;
&lt;details&gt;
&lt;summary&gt;See community builder use cases&lt;/summary&gt;
&lt;ul&gt;
&lt;li&gt;&lt;a href=&#34;https://www.s-anand.net/blog/extracting-ai-advice/&#34;&gt;Extracting AI Advice&lt;/a&gt;: A community manager who has years of recorded Q&amp;amp;A sessions and AMAs can map-reduce all the transcripts to find the top recurring questions members actually ask &amp;ndash; then build a self-serve knowledge base from members&amp;rsquo; own words rather than guessing what to put in it.&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://www.s-anand.net/blog/finding-old-friends-with-gemini/&#34;&gt;Finding Old Friends with Gemini&lt;/a&gt;: A community builder trying to re-engage lapsed members can use Gemini Deep Research to find where they are now &amp;ndash; new employer, new city, new name &amp;ndash; and reach out at the right career moment rather than to a dead email address.&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://www.s-anand.net/blog/using-ai-for-work-news/&#34;&gt;Using AI for Work News&lt;/a&gt;: A community manager can set up a weekly automation that scans public sources for what members have been doing &amp;ndash; new articles, talks, job changes, launches &amp;ndash; and auto-drafts a &amp;ldquo;members in the news&amp;rdquo; section for the newsletter that would otherwise go unwritten for lack of time.&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://www.s-anand.net/blog/transcript-ai-ded-interviews/&#34;&gt;Transcript AI-ded Interviews&lt;/a&gt;: A community builder who runs office hours or mentorship sessions can synthesise all the session transcripts to surface the top recurring problems members raise &amp;ndash; then use that signal to design better programming rather than guessing what the community actually needs.&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://www.s-anand.net/blog/tds-comic-generation/&#34;&gt;TDS Comic Generation&lt;/a&gt;: A community builder can create a recurring comic strip with a few consistent mascots to announce events, explain community norms, or celebrate member milestones &amp;ndash; something memorable and shareable in a way that a plain-text post is not, and producible in minutes with no design budget.&lt;/li&gt;
&lt;/ul&gt;
&lt;/details&gt;
&lt;hr&gt;
&lt;p&gt;When I delivered the &lt;a href=&#34;https://talks.s-anand.net/2025-12-05-scdm-keynote/&#34;&gt;Society for Clinical Data Management keynote&lt;/a&gt;, they audience was surprised how much I knew about their field because I spoke about &lt;a href=&#34;https://talks.s-anand.net/2025-12-05-scdm-keynote/#4&#34;&gt;Informed Consent Forms&lt;/a&gt; and &lt;a href=&#34;https://talks.s-anand.net/2025-12-05-scdm-keynote/#13&#34;&gt;Extracting Schedule of Assessments&lt;/a&gt; and so on. Truth is, I know &lt;em&gt;nothing&lt;/em&gt; about these. Claude created the slides. I asked it to explain enough so I can talk through it.&lt;/p&gt;
&lt;p&gt;I didn&amp;rsquo;t get the implication then, but I think I do now, and the implication is stunning. I now have material to deliver a talk to &lt;em&gt;any&lt;/em&gt; audience.&lt;/p&gt;
&lt;p&gt;So far, I&amp;rsquo;ve been limiting myself to technical talks. Why bother? I can speak to any audience about using AI in their field.&lt;/p&gt;
&lt;ul&gt;
&lt;li&gt;&lt;strong&gt;Human Resources &amp;amp; Organizational Design:&lt;/strong&gt; HR leaders are drowning in qualitative data (interviews, performance reviews, employee sentiment) and are terrified of being left behind by AI.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Marketing &amp;amp; Communications:&lt;/strong&gt; CMOs are under pressure to produce more content with fewer resources. They want to see live workflows of how a single blog post or chat transcript can be repurposed into a full campaign.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Finance &amp;amp; Banking:&lt;/strong&gt; This sector is heavily regulated and drowning in unstructured paperwork. Extract specific clauses from 100-page compliance documents will immediately capture the attention of risk officers.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Event Management (MICE):&lt;/strong&gt; The industry that organizes conferences is itself looking to modernize. Matching attendees, transcribing massive archives of past events, or predict logistical needs is highly relevant.&lt;/li&gt;
&lt;/ul&gt;
&lt;p&gt;Time to get out of my comfort zone!&lt;/p&gt;
&lt;!--
https://chatgpt.com/c/699b13b5-04cc-83a4-b835-9c37dd4ce2cd
https://gemini.google.com/app/f686c3ce03bdb116
--&gt;
</description>
    </item>
    <item>
      <title>When LLM prices fall 10x every year</title>
      <link>https://www.s-anand.net/blog/when-llm-prices-fall-10x-every-year/</link>
      <pubDate>Fri, 20 Feb 2026 09:25:45 +0800</pubDate>
      <guid>https://www.s-anand.net/blog/when-llm-prices-fall-10x-every-year/</guid>
      <description>&lt;!--
https://claude.ai/chat/f0070b78-0653-4172-9906-b6b96b8986dc
https://gemini.google.com/app/939d6b2d87fbe085
https://chatgpt.com/c/6997b88f-f754-83a4-9fa6-362f56c0c3d4
--&gt;
&lt;p&gt;In Feb 2024, Claude 3 Opus was the best model, at $15/MTok.&lt;br&gt;
In Jul 2024, GPT 4o Mini reached that quality at 10% of the price.&lt;br&gt;
In Dec 2024, DeepSeek v3 reached that quality at 1% of the price.&lt;/p&gt;
&lt;video width=&#34;1337&#34; height=&#34;724&#34; style=&#34;max-width: 100%; height: auto;&#34; controls autoplay loop muted&gt;
  &lt;source src=&#34;https://files.s-anand.net/images/2026-02-20-llm-pricing.webm&#34; type=&#34;video/webm&#34;&gt;
  &lt;a href=&#34;https://files.s-anand.net/images/2026-02-20-llm-pricing.webm&#34;&gt;Video&lt;/a&gt;
&lt;/video&gt;
&lt;p&gt;&lt;a href=&#34;https://sanand0.github.io/llmpricing/&#34;&gt;See the interactive version&lt;/a&gt;&lt;/p&gt;
&lt;p&gt;If the price continues to fall 10x every 11-12 months or so (and &lt;a href=&#34;https://claude.ai/share/dd426a79-1bfe-4a5c-94d1-ffc407de6040&#34;&gt;it&lt;/a&gt; &lt;a href=&#34;https://gemini.google.com/share/6c55131f1dcd&#34;&gt;has&lt;/a&gt; &lt;a href=&#34;https://chatgpt.com/share/6997ba52-ff64-8003-aa43-7e8d12818c66&#34;&gt;been&lt;/a&gt;), then in a year, a Claude 4.6 Opus like model will cost 1/10th of the $5/MTok today, and in 2 years, 1/100th of that.&lt;/p&gt;
&lt;p&gt;(We&amp;rsquo;ll be using better models, of course.)&lt;/p&gt;
&lt;p&gt;But 2 years isn&amp;rsquo;t far away. If Opus 4.6 were 100x cheaper, I could do 100x of what I could do with it today. What would we do with it?&lt;/p&gt;
&lt;p&gt;If we assume that they&amp;rsquo;ll become 100x &lt;em&gt;faster&lt;/em&gt; as well (and that&amp;rsquo;s an important assumption), and the reliability will continue to improve, then:&lt;/p&gt;
&lt;ul&gt;
&lt;li&gt;&lt;strong&gt;LLM LSPs&lt;/strong&gt;. &lt;em&gt;Language servers&lt;/em&gt; could be LLMs. Hover over a squiggly line to understand a bug it spotted, right click and fix. Move on.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;LLM pre-commit hooks&lt;/strong&gt;. Write docs, write and run tests, refactor - automatically before you commit.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Continous refactoring&lt;/strong&gt;. LLMs auto-refactor the code, run tests, and commit better code.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Auto-fix from logs&lt;/strong&gt;. Log analysis -&amp;gt; Test case -&amp;gt; Fix -&amp;gt; Deployment can be automated.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Pick best option&lt;/strong&gt;. LLMs generate 30 diverse options for each task, test all, and pick the best. E.g. What&amp;rsquo;s the best language / framework for this? What&amp;rsquo;s the better visual design? What should I build?&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Live docs&lt;/strong&gt;. LLMs auto-update docs every commit.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Adversarial workflows&lt;/strong&gt;. LLMs continously run adversarial test cases to break the code, and fix it.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Build &amp;amp; discard, don&amp;rsquo;t buy&lt;/strong&gt;. Most tools are easier to create than purchase. They&amp;rsquo;re also easier to throw away. To hell with code quality!&lt;/li&gt;
&lt;/ul&gt;
</description>
    </item>
    <item>
      <title>Using AI for work news</title>
      <link>https://www.s-anand.net/blog/using-ai-for-work-news/</link>
      <pubDate>Sat, 14 Feb 2026 14:32:16 +0800</pubDate>
      <guid>https://www.s-anand.net/blog/using-ai-for-work-news/</guid>
      <description>&lt;p&gt;This week, &lt;a href=&#34;https://www.linkedin.com/in/namit-sureka-43ab89&#34;&gt;Namit&lt;/a&gt; and I met a Straive team that operates from a client office. One team member asked:&lt;/p&gt;
&lt;blockquote&gt;
&lt;p&gt;I believe that we are doing wonders out here, but we are closed from what is happening in the rest our organization.&lt;/p&gt;
&lt;p&gt;I want team members to interact with others to see what interesting things they have delivered and where we can implement that solution.&lt;/p&gt;
&lt;p&gt;Could we have sessions, maybe a monthly newsletter, showing what innovations we&amp;rsquo;re working on? This would really keep us engaged with the tech that is going outside of the work that we do.&lt;/p&gt;
&lt;/blockquote&gt;
&lt;p&gt;A good point. This reminded me of an experiment last month.&lt;/p&gt;
&lt;hr&gt;
&lt;p&gt;&lt;a href=&#34;https://studio.workspace.google.com/&#34;&gt;Google Workspace Studio&lt;/a&gt; lets you create automations. For example, here&amp;rsquo;s a &lt;a href=&#34;https://studio.workspace.google.com/workflow/ydef223dcc0e1c583207620711f3fc01a&#34;&gt;flow I set up to create a weekly newsletter about client news&lt;/a&gt;:&lt;/p&gt;
&lt;p&gt;&lt;a href=&#34;https://studio.workspace.google.com/workflow/ydef223dcc0e1c583207620711f3fc01a&#34;&gt;&lt;img loading=&#34;lazy&#34; src=&#34;https://files.s-anand.net/images/2026-02-14-google-workspace-studio.webp&#34;&gt;&lt;/a&gt;&lt;/p&gt;
&lt;ul&gt;
&lt;li&gt;&lt;strong&gt;Step 1: On a schedule&lt;/strong&gt;. Every Monday at 8am&amp;hellip;&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Step 2: Ask Gemini&lt;/strong&gt;. Scan my Google Workspace for the latest client related news and organize it as an email newsletter. Also find the latest news about these clients from the web and put it together seamlessly. Write it in a nice Malcolm Gladwell style narrative.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Step 3: Ask Gemini&lt;/strong&gt;. Just write today&amp;rsquo;s date in YYYY-MM-DD format. Nothing else&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Step 4: Draft an email&lt;/strong&gt;.
&lt;ul&gt;
&lt;li&gt;To: me.&lt;/li&gt;
&lt;li&gt;Subject: Weekly news [Step 3: Content created by Gemini]&lt;/li&gt;
&lt;li&gt;Message: [Step 2: Content created by Gemini]&lt;/li&gt;
&lt;/ul&gt;
&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Step 5: Create a doc&lt;/strong&gt;.
&lt;ul&gt;
&lt;li&gt;New doc name: Weekly news [Step 3: Content created by Gemini]&lt;/li&gt;
&lt;li&gt;Content to add: [Step 2: Content created by Gemini]&lt;/li&gt;
&lt;li&gt;Location for new doc: &lt;a href=&#34;https://drive.google.com/drive/folders/1G3qOPQvtGLVlIeC_7b9SvBmkx52qzg_a&#34;&gt;Weekly news&lt;/a&gt;&lt;/li&gt;
&lt;/ul&gt;
&lt;/li&gt;
&lt;/ul&gt;
&lt;hr&gt;
&lt;p&gt;Now, I get to see a nicely formatted newsletter every Monday morning about new client activity, added to a &lt;a href=&#34;https://drive.google.com/drive/folders/1G3qOPQvtGLVlIeC_7b9SvBmkx52qzg_a&#34;&gt;Weekly news&lt;/a&gt; Google Drive folder accessible to Straive.&lt;/p&gt;
&lt;p&gt;For example, I learnt that:&lt;/p&gt;
&lt;ul&gt;
&lt;li&gt;One client blocked coding agents (Codex, Claude Code, etc) which has slowed down our team.&lt;/li&gt;
&lt;li&gt;Another client needs to modify their API Gateway + Lambda infrastructure to handle long-running agents without exceeding API timeouts&lt;/li&gt;
&lt;li&gt;Yet others requested omni-channel convergence optimization, price sensitivity modeling, payments infrastructure setup, and much more.&lt;/li&gt;
&lt;/ul&gt;
&lt;hr&gt;
&lt;p&gt;I always wished we had an &amp;ldquo;internal reporter&amp;rdquo; at Gramener, going around, interviewing teams, and writing interesting stories about what&amp;rsquo;s happening.&lt;/p&gt;
&lt;p&gt;I did &lt;strong&gt;not&lt;/strong&gt; expect I&amp;rsquo;d be able to hire Malcolm Gladwell (or &lt;em&gt;anyone&lt;/em&gt; I want) as our internal reporter!&lt;/p&gt;
</description>
    </item>
    <item>
      <title>How to Organize Browser Workspaces with LLMs and Data</title>
      <link>https://www.s-anand.net/blog/how-to-organize-browser-workspaces-with-llms-and-data/</link>
      <pubDate>Mon, 07 Apr 2025 04:44:36 +0000</pubDate>
      <guid>https://www.s-anand.net/blog/how-to-organize-browser-workspaces-with-llms-and-data/</guid>
      <description>&lt;p&gt;Here&amp;rsquo;s an example of how I am using LLMs to solve a day-to-day workflow problem.&lt;/p&gt;
&lt;p&gt;Every day, I interact with a barrage of websites: emails, news, social media, and work tools across multiple devices. &lt;a href=&#34;https://learn.microsoft.com/en-us/deployedge/microsoft-edge-workspaces&#34;&gt;Microsoft Edge’s workspaces&lt;/a&gt; syncs groups of websites across devices. I&amp;rsquo;ve never tried it, started today, and wondered: &lt;strong&gt;how should I organize my workspaces?&lt;/strong&gt;&lt;/p&gt;
&lt;div class=&#34;video-embed&#34;&gt;&lt;iframe src=&#34;https://www.youtube.com/embed/1kJ59DzjNOU&#34; title=&#34;YouTube video&#34; loading=&#34;lazy&#34; allow=&#34;accelerometer; autoplay; clipboard-write; encrypted-media; gyroscope; picture-in-picture&#34; allowfullscreen&gt;&lt;/iframe&gt;&lt;/div&gt;
&lt;p&gt;Rather than think (thinking is outdated), I used LLMs.&lt;/p&gt;
&lt;h3 id=&#34;extract-browsing-history&#34;&gt;Extract Browsing History&lt;/h3&gt;
&lt;p&gt;Edge stores website history in a &lt;a href=&#34;https://www.google.com/search?q=Where+is+the+Edge+browser+history+stored+on+Windows+and+Linux%3F&#34;&gt;SQLite database&lt;/a&gt;. But the file is locked by the browser by default. So I spent a fair bit of time figure out how to read it despite it being unlocked. Here are some options:&lt;/p&gt;
&lt;div class=&#34;highlight&#34;&gt;&lt;pre tabindex=&#34;0&#34; class=&#34;chroma&#34;&gt;&lt;code class=&#34;language-bash&#34; data-lang=&#34;bash&#34;&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;datasette .config/microsoft-edge/Default/History --nolock
&lt;/span&gt;&lt;/span&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;sqlite3 &lt;span class=&#34;s1&#34;&gt;&amp;#39;file:.config/microsoft-edge/Default/History?mode=ro&amp;amp;nolock=1&amp;#39;&lt;/span&gt; &lt;span class=&#34;s1&#34;&gt;&amp;#39;SELECT url FROM urls&amp;#39;&lt;/span&gt; &amp;gt; urls.txt
&lt;/span&gt;&lt;/span&gt;&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;p&gt;(&lt;a href=&#34;https://duckdb.org/&#34;&gt;DuckDB&lt;/a&gt; cannot read locked SQLite files - else I&amp;rsquo;d use that.)&lt;/p&gt;
&lt;p&gt;Then comes extracting the hostnames from the URLs. I used &lt;a href=&#34;https://github.com/simonw/llm-cmd&#34;&gt;&lt;code&gt;llm cmd&lt;/code&gt;&lt;/a&gt; to ask Gemini 2.5 Pro:&lt;/p&gt;
&lt;div class=&#34;highlight&#34;&gt;&lt;pre tabindex=&#34;0&#34; class=&#34;chroma&#34;&gt;&lt;code class=&#34;language-bash&#34; data-lang=&#34;bash&#34;&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;llm cmd &lt;span class=&#34;s1&#34;&gt;&amp;#39;Extract just the hostnames from urls.txt which has a list of URLs, one per line. Only pick the https:// URLs. Save into hostnames.txt&amp;#39;&lt;/span&gt;
&lt;/span&gt;&lt;/span&gt;&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;p&gt;I expanded the response &lt;code&gt;awk -F/ &#39;/^https:\/\//{print $3}&#39; urls.txt&lt;/code&gt; into:&lt;/p&gt;
&lt;div class=&#34;highlight&#34;&gt;&lt;pre tabindex=&#34;0&#34; class=&#34;chroma&#34;&gt;&lt;code class=&#34;language-bash&#34; data-lang=&#34;bash&#34;&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;awk -F/ &lt;span class=&#34;s1&#34;&gt;&amp;#39;/^https:\/\//{print $3}&amp;#39;&lt;/span&gt; urls.txt &lt;span class=&#34;p&#34;&gt;|&lt;/span&gt; sort &lt;span class=&#34;p&#34;&gt;|&lt;/span&gt; uniq -c &lt;span class=&#34;p&#34;&gt;|&lt;/span&gt; sort -k 1n &amp;gt; hostnames.txt
&lt;/span&gt;&lt;/span&gt;&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;p&gt;That gave me ~1,400 hostnames.&lt;/p&gt;
&lt;h2 id=&#34;cluster-with-llms&#34;&gt;Cluster with LLMs&lt;/h2&gt;
&lt;p&gt;I passed these to O1 Pro and Gemini 2.5 Pro:&lt;/p&gt;
&lt;blockquote&gt;
&lt;p&gt;Here are the sites I visit, with rough frequency. On Microsoft Edge, I can create workspaces. Based on this browsing behavior, what kinds of workspaces might I create? Give me multiple options.&lt;/p&gt;
&lt;/blockquote&gt;
&lt;p&gt;Both gave a similar set of strategies, which I&amp;rsquo;ve implemented as:&lt;/p&gt;
&lt;ul&gt;
&lt;li&gt;&lt;strong&gt;Main&lt;/strong&gt;: email, calendar, tasks, etc.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Work&lt;/strong&gt;: work related sites (drive, expenses, HR platform, etc.)&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Chill&lt;/strong&gt;: YouTube, Minesweeper, Netflix, etc.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Read&lt;/strong&gt;: blogs, articles, stuff I need to catch-up on&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Code&lt;/strong&gt;: GitHub, StackOverflow, CodePen, etc.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Chores&lt;/strong&gt;: government services, shopping, etc.&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;AI&lt;/strong&gt;: ChatGPT, Gemini, Perplexity, etc.&lt;/li&gt;
&lt;/ul&gt;
&lt;p&gt;I was surprised how &lt;strong&gt;similar&lt;/strong&gt; a strategy both models converted to. Either these models &lt;strong&gt;really&lt;/strong&gt; think alike, or my browsing pattern is a fairly common one. (My guess is the latter.)&lt;/p&gt;
&lt;h3 id=&#34;write-with-llms&#34;&gt;Write with LLMs&lt;/h3&gt;
&lt;p&gt;After setting up my groups, I needed to write this post. Instead of slow typing, I stepped out and &lt;a href=&#34;https://chatgpt.com/share/67f36093-e810-800c-a9d0-de9bfb7ecf86&#34;&gt;talked with ChatGPT&lt;/a&gt;. (Talking to a machine in the office felt strange, so I changed my space.) I explained my whole process, and in about eight minutes, the first draft was done. Normally, writing takes much longer, but the voice chat made it quick and smooth.&lt;/p&gt;
&lt;p&gt;The editing after that was manual and took 20 minutes.&lt;/p&gt;
&lt;h3 id=&#34;things-i-learnt&#34;&gt;Things I learnt&lt;/h3&gt;
&lt;ul&gt;
&lt;li&gt;&lt;strong&gt;Simple Patterns&lt;/strong&gt;: My browsing history shows clear patterns. AI helped me find groups I couldn&amp;rsquo;t see before&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Small Fixes - Big Wins&lt;/strong&gt;: A small challenge (opening a locked file) taught me a bunch of new useful stuff&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Voice Made It Easy&lt;/strong&gt;: Talking with ChatGPT made writing fast and easy. It shows that speaking to a machine can save time&lt;/li&gt;
&lt;/ul&gt;
</description>
    </item>
    <item>
      <title>Things I Learned - 10 Nov 2024</title>
      <link>https://www.s-anand.net/blog/things-i-learned-10-nov-2024/</link>
      <pubDate>Sun, 10 Nov 2024 00:00:00 +0000</pubDate>
      <guid>https://www.s-anand.net/blog/things-i-learned-10-nov-2024/</guid>
      <description>&lt;p&gt;This week, I learned:&lt;/p&gt;
&lt;ul&gt;
&lt;li&gt;&lt;a href=&#34;https://openfreemap.org/&#34;&gt;OpenFreeMap&lt;/a&gt; is a free embeddable OpenStreetMap tile server. You can use &lt;a href=&#34;https://maplibre.org/&#34;&gt;MapLibre GL&lt;/a&gt; (more features) or Leaflet (simpler) to render it. It offers styling and self-hosting.&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://actions.zapier.com/&#34;&gt;Zapier Actions&lt;/a&gt; are an easy way to set up custom actions like GMail / Google Calendar APIs for GPTs, since &lt;a href=&#34;https://community.openai.com/t/gpt-oauth-callback-url-keeps-changing/493236&#34;&gt;GPTs&amp;rsquo; callback URLs keep changing&lt;/a&gt;. But they fail often, and don&amp;rsquo;t work on mobile. At least for me.&lt;/li&gt;
&lt;li&gt;LLM Vision Use Cases in manufacturing and earth sciences (via Shivku)
&lt;ul&gt;
&lt;li&gt;Automated geoscience image descriptions &lt;a href=&#34;https://www.linkedin.com/posts/paulhcleverley_geosciences-earthscience-geology-activity-7254037937674240000-pQab/&#34;&gt;Ref&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;Interpret Wind Turbine photos and charts, construction monitoring, equipment maintenance &amp;amp; charts &lt;a href=&#34;https://www.linkedin.com/pulse/vision-ai-energy-use-cases-copilot-wind-siting-impact-kalyanaraman-wqe7c/&#34;&gt;Ref&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;Forecast weather based on cloud photos! &lt;a href=&#34;https://www.linkedin.com/pulse/cloud-typing-local-weather-forecasting-using-chatgpt-cam-shivkumar-1hhkc/&#34;&gt;Ref&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;Analyze thermal image of solar panels, electroluminescence images for warranty claims, ROI estimates from Google Sunroof rooftop images &lt;a href=&#34;https://www.linkedin.com/pulse/vision-ai-energy-use-cases-part-1-copilot-solar-pv-kalyanaraman-ccszc/&#34;&gt;Ref&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;Corrosion detection in electricity towers, turbines, storage tanks, penstock. Interpret non-destructive test images &lt;a href=&#34;https://www.linkedin.com/pulse/vision-ai-energy-use-cases-copilot-corrosion-shivkumar-kalyanaraman-onuic/&#34;&gt;Ref&lt;/a&gt;&lt;/li&gt;
&lt;/ul&gt;
&lt;/li&gt;
&lt;li&gt;Google counts auto-completion when saying &amp;ldquo;25% of all the code is written by AI at Google&amp;rdquo;. &amp;ldquo;It&amp;rsquo;s a helpful productivity tool but it&amp;rsquo;s not doing any engineering at all. It&amp;rsquo;s probably about as good, maybe slightly worse, than Copilot.&amp;rdquo; &lt;a href=&#34;https://news.ycombinator.com/item?id=42002212&#34;&gt;YCombinator&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;Workflow for AI video creation: Use Meshcapade (meshcapade.com) to generate body movement of a 3D-rendered character. Pass that video to Runway&amp;rsquo;s video-to-video model to generate any visual. Add music from Suno &lt;a href=&#34;https://www.linkedin.com/posts/peter-gostev_i-discovered-a-really-cool-new-workflow-for-activity-7260003053771141120-DJpS&#34;&gt;Ref&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;Someone sorted the X and Y columns independently for regression. &lt;a href=&#34;https://stats.stackexchange.com/q/185507&#34;&gt;Ref&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;Android keyboard learning only sends model changes back to server and not local keywords. Model changes are aggregated! &lt;a href=&#34;https://chatgpt.com/share/672d6d6d-46a0-800c-a130-c689f5ebc0b7&#34;&gt;Ref&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;Here is a prompt for audio transcription using Gemini. &lt;a href=&#34;https://gist.github.com/rajivsinclair/8fb0371f6eda25f9e5cc515cd77abd62&#34;&gt;Ref&lt;/a&gt;
&lt;ul&gt;
&lt;li&gt;Transcription: Accurately transcribe the audio clip in the original language. Include all spoken words, fillers, slang, colloquialisms, and any code-switching instances. Pay attention to dialects and regional variations common among immigrant communities. Do your best to capture the speech accurately, and flag any unintelligible portions with &lt;code&gt;[inaudible]&lt;/code&gt;.&lt;/li&gt;
&lt;li&gt;Translation: Translate the transcription into English. Preserve the original meaning, context, idiomatic expressions, and cultural references. Ensure that nuances and subtleties are accurately conveyed.&lt;/li&gt;
&lt;li&gt;Capture Vocal Nuances: Note vocal cues such as tone, pitch, pacing, emphasis, and emotional expressions that may influence the message. These cues are critical for understanding intent and potential impact.&lt;/li&gt;
&lt;/ul&gt;
&lt;/li&gt;
&lt;li&gt;Here are some approaches to large-scale classification of medical codes. &lt;a href=&#34;https://chatgpt.com/share/672dd476-7694-800c-a150-f3de912788ef&#34;&gt;ChatGPT&lt;/a&gt;
&lt;ul&gt;
&lt;li&gt;Fine-Tuning LLMs on Medical Data: Enhance LLMs by training them on medical datasets, such as clinical notes and discharge summaries, to improve their understanding of medical terminology and context.&lt;/li&gt;
&lt;li&gt;Multi-Agent Frameworks: Implement a multi-agent system that simulates real-world coding processes with distinct roles (e.g., patient, physician, coder, reviewer, adjuster). Each agent utilizes an LLM to perform specific functions, enhancing interpretability and reliability. &lt;a href=&#34;https://arxiv.org/abs/2406.15363&#34;&gt;ArXiv&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;Retrieve-Rank Systems: Develop a two-stage system where the LLM first retrieves potential ICD-10 codes and then ranks them based on relevance, improving precision in code assignment. &lt;a href=&#34;https://arxiv.org/abs/2407.12849&#34;&gt;ArXiv&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;Embedding-Based Approaches: Use LLMs to generate embeddings for ICD-10 codes and medical texts, facilitating the matching of texts to appropriate codes through similarity measures. &lt;a href=&#34;https://github.com/kaneplusplus/icd-10-cm-embedding&#34;&gt;GitHub&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;Hierarchical Classification: Leverage the hierarchical structure of ICD-10 codes by first classifying texts into broader categories before assigning specific codes, reducing complexity and improving accuracy. &lt;a href=&#34;https://arxiv.org/abs/2310.06552&#34;&gt;ArXiv&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;Two-Stage Verification Models: Combine LLMs with verification models, such as Long Short-Term Memory (LSTM) networks, to validate and refine the codes suggested by the LLM, balancing recall and precision. &lt;a href=&#34;https://arxiv.org/abs/2311.13735&#34;&gt;ArXiv&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;Also, a mixture of models approach might work. Feed any existing NLP model / rules as a second opinion.&lt;/li&gt;
&lt;/ul&gt;
&lt;/li&gt;
&lt;li&gt;GraphRAG is better if data is naturally graph-structured. Else, it&amp;rsquo;s slow and fills up the context window with even vaguely related stuff. Vigneshbabu, AMAT.&lt;/li&gt;
&lt;li&gt;ChatGPT for Windows desktop supports real-time voice and a global shortcut (Alt Space).&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://uithub.com&#34;&gt;uithub&lt;/a&gt; converts GitHub repos to Markdown. Just replace &amp;ldquo;g&amp;rdquo; in &amp;ldquo;github.com/&amp;hellip;&amp;rdquo; with &amp;ldquo;u&amp;rdquo;. &lt;a href=&#34;https://uithub.com/gramener/asyncllm&#34;&gt;Example&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;WebContainers are a thing and Bolt.new uses them!&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://github.com/DS4SD/docling&#34;&gt;Docling&lt;/a&gt; by IBM converts PDF, DOCX, etc. to Markdown. Like &lt;a href=&#34;https://pymupdf.readthedocs.io/en/latest/pymupdf4llm/&#34;&gt;PyMuPDF4LLM&lt;/a&gt; but better.&lt;/li&gt;
&lt;li&gt;Check out &lt;a href=&#34;https://www.loom.com/&#34;&gt;Loom&lt;/a&gt; and &lt;a href=&#34;https://cleanshot.com/&#34;&gt;Cleanshot&lt;/a&gt; are the recommended tools for screen recording and screenshotting. But Loom is paid and Cleanshot is Mac only.&lt;/li&gt;
&lt;li&gt;The Rubik&amp;rsquo;s cube has a Hamiltonian cycle through every one of its 43 quintillion states. &lt;a href=&#34;https://bruce.cubing.net/ham333/rubikhamiltonexplanation.html&#34;&gt;Ref&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://microsoft.github.io/OmniParser/&#34;&gt;OmniParser&lt;/a&gt; is great at parsing screenshots and identifying bounding boxes.&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://www.recraft.ai/&#34;&gt;Recraft.ai&lt;/a&gt; is currently SOTA in text to image. It&amp;rsquo;s fairly impressive and could be a good alternative to Figma.&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://zed.dev/&#34;&gt;Zed.dev&lt;/a&gt; is an AI code editor by the creators of Atom. It&amp;rsquo;s written in Rust and is blazing fast. It has native AI integration.&lt;/li&gt;
&lt;li&gt;Artificial Analysis has a bunch of new leaderboards and arenas.
&lt;ul&gt;
&lt;li&gt;Open AI TTS leads the &lt;a href=&#34;https://artificialanalysis.ai/text-to-speech/arena?tab=Leaderboard&#34;&gt;TTS Leaderboard&lt;/a&gt;. ElevenLabs is a bit behind.&lt;/li&gt;
&lt;li&gt;Recraft V3 &amp;gt; Flux 1.1 leads &lt;a href=&#34;https://artificialanalysis.ai/text-to-image/arena?tab=Leaderboard&#34;&gt;Text to Image Leaderboard&lt;/a&gt;&lt;/li&gt;
&lt;/ul&gt;
&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://github.com/Standard-Intelligence/hertz-dev&#34;&gt;Hertz-Dev&lt;/a&gt; is an open source realtime voice chat model. But it doesn&amp;rsquo;t fit in Google Colab T4&amp;rsquo;s RAM&lt;/li&gt;
&lt;li&gt;Chain of Thought reduces performance where thinking makes humans worse. &lt;a href=&#34;https://arxiv.org/abs/2410.21333&#34;&gt;Ref&lt;/a&gt;. Specifically:
&lt;ul&gt;
&lt;li&gt;Artificial grammar learning&lt;/li&gt;
&lt;li&gt;Facial recognition&lt;/li&gt;
&lt;li&gt;Classifying data that has exceptions&lt;/li&gt;
&lt;/ul&gt;
&lt;/li&gt;
&lt;li&gt;&lt;a href=&#34;https://hamel.dev/blog/posts/llm-judge/&#34;&gt;Creating a LLM-as-a-Judge That Drives Business Results&lt;/a&gt; by Hamel Husain.
&lt;ul&gt;
&lt;li&gt;Get THE domain expert (or approver) as the tester.&lt;/li&gt;
&lt;li&gt;Create a dataset that is DIVERSE.&lt;/li&gt;
&lt;li&gt;Covers EACH combination of:
&lt;ul&gt;
&lt;li&gt;Features&lt;/li&gt;
&lt;li&gt;Scenarios: e.g. multiple matches, no match, ambiguous request, invalid/incomplete input, unsupported feature, system error&lt;/li&gt;
&lt;li&gt;Persona: e.g. new user, expert user, non-native speaker, busy professional, technophobe, elderly user&lt;/li&gt;
&lt;/ul&gt;
&lt;/li&gt;
&lt;li&gt;Generate data using existing data + synthetic data for each SPECIFIC combination of the above&lt;/li&gt;
&lt;li&gt;Evaluate based only on PASS/FAIL with a CRITIQUE detailed enough for a new employee. Include:
&lt;ul&gt;
&lt;li&gt;Nuances: Something a failed response did well or a passed response didn&amp;rsquo;t quite do well&lt;/li&gt;
&lt;li&gt;Improvements: Suggest how model can improve&lt;/li&gt;
&lt;/ul&gt;
&lt;/li&gt;
&lt;li&gt;Build an SPA to make it easy for the domain expert to review&lt;/li&gt;
&lt;/ul&gt;
&lt;/li&gt;
&lt;li&gt;LLMs can be made to unlearn (copyright material) better by identifying components related to the knowledge to unlearn and applying a larger learning rate to these while leaving other parts unchanged. As opposed to low learning rates for all components. &lt;a href=&#34;https://arxiv.org/abs/2410.16454&#34;&gt;Ref&lt;/a&gt;&lt;/li&gt;
&lt;/ul&gt;
</description>
    </item>
  </channel>
</rss>
