← Back to context

Comment by ModernMech

3 hours ago

Again it's just not been my experience through testing so I'm curious what kind of measurements you're citing here.

Here's some related work that matches my observations: https://danluu.com/pl-tokens/

  • Thanks, yeah I remember when that made the rounds a couple weeks ago. Although what I'm proposing is a little different than what's covered there: the AI designing a language for a particular task, writing the runtime to implement the language, and then solving the task in the language it designed. The blog rather is about how an AI performs with languages designed by people for general purposes.

    • What evals did you use to compare that to using an existing language with a large corpus of training data?