Post #4257284
2026-07-27 07:46 UTC
Am I in The Stack V3, https://huggingface.co/spaces/HuggingFaceCode/in-the-stack.
> The Stack v3 is a 15.9 TB dataset of source code across 713 programming languages from 173M repositories, crawled from GitHub in 2025
ALL my projects have been used to train this model. All of them. Including orgs I contribute to.
…
What if an LLM generates code from my PhD thesis, or any algorithms I wrote? Good luck with the copyrights.
#llm #ai
Replies (0)
No replies.