Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

What kind of data or data formatting do you need to fine-tune GPT-4o? I'd love to throw a bunch of documentation at it and let it learn, but I don't have the resources to extract knowledge, format it as questions and answers, etc.


Fine-tuning generally isn't an effective way to add extra knowledge from things documentation - my understanding is that the vast amounts of knowledge in the original training data tend to overwhelm any extra knowledge you try to add by fine-tuning.

OpenAI's documentation has good examples of how the data should be formatted (and when it's appropriate to fine-tune): https://platform.openai.com/docs/guides/fine-tuning/fine-tun...


does the same result apply to LoRA?


Fine tuning is most often performed by training a LoRA. It's almost certainly what OAI are doing, as they can inference many lightweight LoRA in parallel atop the same foundation model.





Consider applying for YC's Fall 2026 batch! Applications are open till July 27.

Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: