Logs

I've been digging more into evals. I wrote a simple Claude completion function in openai/evals to better understand how the different pieces fit together. Quick and dirty code:

I can't believe I am saying this but if you play around with language models locally, a 1 TB drive, might not be big enough for very long.

As someone learning to draw, I really enjoyed this article: <https://maggieappleton.com/still-cant-draw>. I've watched the first three videos in this playlist so far and have been sketching random...

One of the greatest misconceptions concerning LLMs is the idea that they are easy to use. They really aren’t: getting great results out of them requires a great deal of experience and hard-fought...

Did a bit more work on a LLM evaluator for connections. I'm mostly trying it with gpt-4 and claude-3-opus. On today's puzzle, the best either did was 2/4 correct. I'm unsure how much more improvement...

I use the <kbd>hyper</kbd>+<kbd>u</kbd> keyboard shortcut to open a language model playground for convenience. I might use this 10-20 times a day. For the last year or so that I've been doing this,...