GitHub will train Copilot models on user interaction data — unless you opt out
Starting April 24, GitHub will begin using interaction data from Copilot Free, Pro, and Pro+ users to train and improve its AI models. The change covers inputs, outputs, code snippets, and associated context. Copilot Business and Copilot Enterprise users are excluded from this update.
Users who don't want their data used for training can opt out in GitHub's settings under "Privacy." If you previously disabled the setting for GitHub to collect this data for product improvements, that preference carries forward — your data won't be used for training unless you explicitly opt in.
What interaction data is collected
The program draws on real-world usage patterns to improve model performance. The data GitHub may collect includes:
- Outputs accepted or modified by you
- Inputs sent to GitHub Copilot, including code snippets shown to the model
- Code context surrounding your cursor position
- Comments and documentation you write
- File names, repository structure, and navigation patterns
- Interactions with Copilot features (chat, inline suggestions, etc.)
- Your feedback on suggestions (thumbs up/down ratings)
The program does not use interaction data from Copilot Business, Copilot Enterprise, or enterprise-owned repositories, nor from users who opt out in their Copilot settings.
GitHub also clarifies that content from your issues, discussions, or private repositories is not used "at rest." Copilot does process code from private repositories while you're actively using it — that data is necessary to run the service and could be used for model training unless you opt out.
Data sharing and model improvements
Data collected under this program may be shared with GitHub affiliates, including Microsoft, but won't be shared with third-party AI model providers or other independent service providers.
The move follows a period of incorporating interaction data from Microsoft employees, which GitHub says led to meaningful improvements, including increased acceptance rates in multiple languages. GitHub plans to begin using interaction data from its own employees as well.
Users with questions can visit GitHub's FAQ and related discussion.



