Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

One question, OP, how does cost for this work? Do you pay the LLM inference cost (quite literally if using an external API) every time you run a line that involves natural language computation? E.g. what happens if you call a "symbolic" function in a loop.


Yes, that's correct. If using say openai, then every semantic ops are API calls to openai. If you're hosting a local LLM via llama.cpp, then obviously there's no inference cost other than that of hosting the model.


This will need a cache of some sort




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: