Scale up as you grow — whether you're running one virtual machine or ten thousand.

From GPU-powered inference and Kubernetes to managed databases and storage, get everything you need to build, scale, and deploy intelligent applications.

This textbox defaults to using Markdown to format your answer.
You can type !ref in this text area to quickly search our full set of tutorials, documentation & marketplace offerings and insert the link!
These answers are provided by our Community. If you find them useful, show some love by clicking the heart. If you run into issues leave a comment, or add your own answer to help others.
Hi there,
Really great to see you pushing the GenAI platform this way, you’re getting into the kind of real-world use cases that can really help shape future improvements!
From what I understand, with a single large .csv, it’’s tricky to track individual sources properly since everything points back to the same file. Splitting into multiple .csv files (one per logical source) might be the more reliable approach right now, similar to how multiple .md files work. But I’m not 100% sure if that’s the only way, it might be worth checking directly with DigitalOcean Support.
Also, full support for .parquet files or more flexible metadata would definitely be a great improvement.
I’d really encourage you to send this feedback to DigitalOcean Support, you’re raising exactly the kinds of points that could help improve the product and documentation over time.
- Bobby