GitHub Copilot updates data usage policy for model training

Policy change details
GitHub announced that from April 24, 2026 onward, interaction data from Copilot Free, Pro, and Pro+ users will be used to train and improve their AI models unless users opt out. Copilot Business and Copilot Enterprise users are not affected by this update.
If you previously opted out of data collection for product improvements, your preference has been retained. You can opt out in settings under "Privacy."
What data is collected
The interaction data that may be collected and leveraged includes:
- Outputs accepted or modified by you
- Inputs sent to GitHub Copilot, including code snippets shown to the model
- Code context surrounding your cursor position
- Comments and documentation you write
- File names, repository structure, and navigation patterns
- Interactions with Copilot features (chat, inline suggestions, etc.)
- Your feedback on suggestions (thumbs up/down ratings)
What data is NOT used
This program does not use:
- Interaction data from Copilot Business, Copilot Enterprise, or enterprise-owned repositories
- Interaction data from users who opt out of model training in their Copilot settings
- Content from your issues, discussions, or private repositories at rest
GitHub notes they use the phrase "at rest" deliberately because Copilot does process code from private repositories when you are actively using Copilot. This interaction data is required to run the service and could be used for model training unless you opt out.
Data sharing and background
The data used in this program may be shared with GitHub affiliates, including Microsoft. This data will not be shared with third-party AI model providers or other independent service providers.
GitHub states they've already been incorporating interaction data from Microsoft employees and have seen meaningful improvements, including increased acceptance rates in multiple languages. They will also begin using interaction data from GitHub employees.
GitHub's initial models were built using a mix of publicly available data and hand-crafted code samples.
📖 Read the full source: HN LLM Tools
👀 See Also

Illinois Passes SB 315: Third-Party Audits Required for Frontier AI Labs
Illinois passes SB 315 requiring frontier AI labs like OpenAI, Anthropic, and Google DeepMind to have safety practices audited by independent third parties. If signed, it becomes the strongest US state AI safety law.

AI Tools May Lead to Homogenized Output in Creative and Development Work
A Reddit user reports that multiple teams using AI tools like ChatGPT, Co-Pilot, and Claude for strategy roadmaps and software development are producing similar outputs with identical buzzword patterns and design structures.

Reddit discussion highlights shift from chatbots to autonomous agents with local execution
A Reddit post distinguishes chatbots from autonomous agents using concrete examples and notes the trend toward local execution with models like LLaMA running on private workstations.

Hacker News AI Discussion Shifts from Demos to Tooling Focus
Recent Hacker News discussions about AI are moving from one-off demos to durable tools like price tracking, verification, memory, evaluation, and workflow integration. This signals a shift toward operationalization where communities stop rewarding novelty-first posts.