Loved this Rocky Aur Rani Kii Prem Kahaani scene where Ranveer asks, “Chinese ko Chinese bol sakte hai?” हम बहनदी भी नहीं बोल सकते? आंटी, मैं दिल्ली से हूँ। मैं कैसे नहीं बहनदी बोलूं बहनदी!? कैसा जमाना आ गया है? फैट-ों को फैट नहीं बोल सकते, ब्लैक-ों को ब्लैक नहीं बोल सकते, ओल्ड-ों को ओल्ड नहीं बोल सकते, मुँह खोलने से डर लगता है मुझे! आप मुझे बताओ, चाइनीज़ को चाइनीज़ बोल सकते हैं? ...

Things I Learned - 21 Jul 2024

This week, I learned: GPT For Work has a set of useful spreadsheet LLM functions Xata offers a free PostgreSQL tier with REST API Mamba now uses mambaforge as the default installation, i.e. conda-forge is the default and only channel! Update: 6 Jun 2025. Mambaforge is sunset as of 29 Jul 2024. Conda-forge now uses Miniforge as the standard installer Ref conda-forge.org. Users should switch to Miniforge instead. nginx supports a load-balancing method least_conn which is far better than the default round-robin. #IMPOSSIBLE LLMs cannot provide a bounding box of objects in images. (Maybe Florence 2 can). Update: Mar 2025. Gemini has good timestamps and bounding boxes Models gently grow in capability. It helps to maintain an impossibility list that steadily gets invalidated. Ref Github Copilot internals walks through how Copilot constructs its prompts

Things I Learned - 14 Jul 2024

This week, I learned: Carlton’s TDS session Always create a new venv via VS Code when starting a training session. Helps reproduce issues (though I could use Colab instead) Create an empty .ipynb notebook and double-click it. That’s another way (though slower) to open a Jupyter notebook Share Parrish Knowledge Project podcast. Three generations of wealth There is a big difference between liking animals and being a vet. Between liking education and being a teacher. Even if no one reads your writing, you benefit from the writing. Emotional.crises like 9/11 or Covid are far easier for markets to recover from Hidden brain podcast. White trying to hard can back fire on you Sometimes conscious thinking makes our automated responses of sports music, dance are great examples Instead, SURRENDER to something outside of you. Like playing with kids. Exercise also sends blood away from brain. Drugs. ChatGPT. It’s called Ue in Chinese philosophy A quick check on the pricing of text to speech models OpenAI TTS: $15/1M chars Ref Deepgram Aura: $15/1M chars Ref Elevenlabs Scale: $165/1M chars Ref Google TTS Neural2: $16/1M chars Ref Azure AI Speech: $15/1M chars Ref AWS Polly Neural TTS: $16/1M chars Ref

I'll leave tomorrow's problems to tomorrow's me

What a delightful idea. I’ll leave tomorrow’s problems to tomorrow’s me. – Saitama, One Punch Man Saitama is now one of my favorite heroes. Right up there with Atticus Finch and Juror #8. Very few people can articulate such a wonderful philosophy as effectively. The closest was Calvin. Of course, it’s not a perfect system. But they do say, “Sometimes, the best way to get something is to stop trying to get it.”

Things I Learned - 07 Jul 2024

This week, I learned: Predibase uses LORAX to run multiple fine-tunings of a base model in a single GPU via adapters. Ref

Things I Learned - 30 Jun 2024

This week, I learned: Amara’s law: “We tend to overestimate the effect of a technology in the short run and underestimate the effect in the long run.” LLM Patterns include Evals, RAG, Fine-tuning, Caching, Guardrails, Defensive UX, Collect feedback. Notably: Defensive UX: Microsoft, Google, and Apple have guidelines for Human-AI interactions Collect feedback: Explicit and implicit Rouge and Context Precision are metrics to evaluate LLM responses that serve as a starting point – but not sufficient, usually Any word with the letters izehsglbo can be spelt on a calculator. That includes Hobbes (538804)! Via Calculator spelling Tor Browser + DuckDuckGo is good for torrent searches. Maybe the Dark Web IS the original Internet. The ad-free hacker web

Hobbes on a calculator

I just learned that any word made of just these letters beighlosz can be spelt on a calculator. That includes Hobbes! 538804 upside-down looks like this: I’m surprised I never knew that. The longest, by far, appears to be hillbillies – 53177187714

Things I Learned - 23 Jun 2024

This week, I learned: Luma Labs Dream Machine generated videos. It’s free and is of reasonable quality. Update: 6 Jun 2025. Costs $10/month LLM DataHub has LLM training datasets, regularly updated From Dan Becker on running a workshop Answer questions at the end, not in parallel in a chat, to avoid distraction Have fewer words in slides when presenting. It’s less distracting Morgan Housel Shane Parrish podcast Risk is what stops you from achieving YOUR goals. What’s risky for me may not be risky for you The lesson from compounding is that you want to optimize for duration, not return. That’s what does the heavy lifting. Survival, consistency, long term - these matter. The performance does NOT matter.

The psychology of peer reviews

We asked the ~500 students in my Tools in Data Science course in Jan 2024 to create data visualizations. They then evaluated each others’ work. Each person’s work was evaluated by 3 peers. The evaluation was on 3 criteria: Insight, Visual Clarity, and Accuracy (with clear details on how to evaluate.) I was curious to see if what we can learn about student personas from their evaluations. ...

Embeddings in DuckDB

This article on Using DuckDB for Embeddings and Vector Search by Sören Brunk shows a number of DuckDB features I wasn’t aware of. DuckDB can read directly from Huggingface datasets DuckDB can read just the parts of a .parquet file it needs, even over HTTP DuckDB lets you write custom functions in Python DuckDB now has a vector similarity search extension I’ve recently become a DuckDB fan and continue to be impressed.

Things I Learned - 09 Jun 2024

This week, I learned: httpretty can mock ALL Python HTTP libraries Japanese pray to dead parents instead of gods. The dead are preserved in plates by priests. Japanese are generally non religious Looks like GPT-4o is using CNNs to create vector embeddings of images, with images gridded into a 1x1, 2x2, etc. PLUS OCR. Ref The sum of a sinusoidal series is like a spirogram. Spinning circle linked to another and so on https://www.andreinc.net/2024/04/24/from-the-circle-to-epicycles

Things I Learned - 02 Jun 2024

This week, I learned: Modal.com seems of offer reasonably priced GPUs Combining vector search and keyword search with reciprocal rank fusion seems to work well for RAG. Ref Knowledge Project podcast. Morgan Housel Differences of opinion exist because of different stories arising from origins and experiences. We are not debating facts. We are debating life lessons! Solution: hear their anecdotes. The stories that taught them their lessons. AI reporting templates are a trend. Domain expertise comes in via structuring the report template and associated prompts. Some audio embedding models: unoti/voice-embeddings, retkowsky/audio_embeddings, pyannote/embedding (for speaker similarity), and more. Hidden Brain podcast: Innovation 2.0: The power of less Subtraction is hard because we are biologically and economically wired against it. It’s also hard because there are fewer markers of subtraction. Additions are natural markers / triggers. Marie Kondo suggests keeping only what sparks joy #POST I tried Undermind.ai - an agent that researches for you. It guides you to ask a detailed question, spends 2-3 minutes finding the answer, and provides detailed results. But it’s worth the wait. It’s a good alternative to quick validations on SciSpace. For popular results, search actually makes results worse! When not to trust language models Perception of fluency and usefulness are NEGATIVELY correlated in LLM! Evaluating Verifiability in Generative Search Engines GPTs are now available to non paying users. Apparently for a few weeks! Everyone also has limited access to GPT-4o. Discussion with Anand Explore BBC Microbit Everyone should get a Raspberry Pi! Watch 2 minutes paper on YouTube More LLM routers: LiteLLM: Open source, OpenAI compatible, 100+ LLMs RouteLLM: Open source, OpenAI compatible, automatically routes based on cost OpenRouter: OpenAI compatible API, several models Unify: Supports many models Portkey: Supports popular providers Martian: Limited set of models d-id and Heygen can modify videos of a person.

Things I Learned - 26 May 2024

This week, I learned: My home WiFi is on WiFi 6. This supports beam-forming which increases range by “focusing” on devices! Predibase lets you run fine-tuned models at the same price, on a per-token basis. 25c/MTok up to 21B models. That’s sames as Claude 3 Haiku, but with fine-tuning. RunPod’s vLLM endpoint lets you run any HuggingFace LLM with an OpenAI API priced on usage (serverless) not on idle time. “Autoscaling to 0”. Portkey is an LLM router

Things I Learned - 19 May 2024

This week, I learned: In Scandinavia, Århus comes after Zürich because Å is a different letter. It was added by the Dutch after WW2 to distance themselves from the Germans. via Zalgo text is where we combine multiple Unicode combining characters Artificial Analysis benchmarks LLM APIs on speed, cost, and quality.

Things I Learned - 12 May 2024

This week, I learned: Radio free Xp podcast. Nudge 61 always announce first before doing. Give people time to plan comment and react. That gets you alignment without sacrificing freedom. give information, not orders. When someone is parking a car, tell them how much space they have, don’t tell them to start stop or how much to turn left it’s almost impossible to change the culture if you’re not the boss

There are 4 frontier #LLMs today. No other (popular) model beats them on BOTH cost and quality. llama-3-8b-instruct claude-3-haiku-20240307 llama-3-70b-instruct gpt-4o-2024-05-13 This list changes rapidly. But in practice, it means there’s little reason to use any other LLM. They beat every other model on cost and quality (measured by the LMSYS Arena ELO score.) I opened Straive + Gramener’s keynote yesterday at marcus evans Group’s Digitech forum with this. Strange that this is not well known. Especially as switching from GPT-4 to Claude 3 Haiku can shrink a $1.2 million Gen AI budget to just $10K. ...

250 BC is when I’d pick to time-travel to. Ashoka was turning into one of the most famous emperors of India and Archimedes was growing into one of the greatest mathematicians of all time. Parallel Lives is a beautiful visualization by Jan Willem Tulp that shows who lived when, showing overlaps, and sized by their prevalence on Wikipedia. I’m a history fan and have spent several hours scrolling through the site: ...

Things I Learned - 05 May 2024

This week, I learned: Hidden brain podcast. Innovation 2.0 solve your own problem. Don’t solve other people’s problems. This helps you pick what you’re good at affordable losses. Make sure you survive borrow others’ spares. spare time, scrap data, anything others don’t use. If you can monetize it, you can pay them back focus on the controllable. Ignore what’s outside your control don’t even waste time on it curl supports globbing, emails Beetrove is a ranking of the popularit of OpenAI GPTs Gemini Prompt Guide has detailed examples of how each role can use Gemini ESLint’s new flat configuration does not support package.json

Things I Learned - 28 Apr 2024

This week, I learned: Tough prompt to test: Gr brx vshdn Fdhvdu flskhu? is a quick way to assess LLM capability. Ref Cheap cloud GPU services thread on Twitter lists: Runpod (17) Vast.ai (17) Modal Labs (8) fly.io (4) LightningAI (4) Colab (4) AkashNet (4) Lambda Labs (4) ShadeFormAI (3) Mac Mini (3) Tensor Dock (2) Hetzner (2) BrevDev (2) JSR lets you publish Deno packages that can be imported by npm via. It also auto-evaluates documentation and scores it! via Snowflake Arctic Cookbook explains how mixture of experts models work A long list of LLM courses online Embeddings can be averaged. So, to embed large documents, average the embeddings of their chunks! OpenAI suggests this.

A quick way to assess LLM capabilities

Simon Willison initiated this very interesting Twitter thread that asks, “What prompt can instantly tell us how good an LLM model is?” The Sally-Anne Test is a popular test that asks: Sally hides a marble in her basket and leaves the room. While she is away, Anne moves the marble from Sally’s basket to her own box. When Sally returns, where will she look for her marble?" ...