Danish surveys linked to administrative records show task reorganization around content generation, AI oversight and integration. Difference-in-differences estimates find no average earnings or recorded-hours effects larger than 2% within two years of ChatGPT’s launch. These early, context-specific results do not imply that future labor-market impacts will remain small.
The report describes AI use through task complexity, skills, purpose, autonomy and success, incorporating success rates into efficiency estimates. Tasks within one occupation can differ substantially, so usage counts alone cannot measure automation. Many metrics rely on model assessment of platform conversations rather than causal measurement of economy-wide productivity.
DeepMind presents Gemini Robotics 2, combining visual-language understanding with whole-body control, dexterity and coordination. Demonstrations extend AI assistance into physical environments. Evidence mainly comes from developer demonstrations and evaluations; it does not establish universal reliability, safety or commercial returns in uncontrolled workplaces.
GenCast produces ensembles of possible weather trajectories up to 15 days ahead. Retrospective evaluations showed improvements over the comparator on many weather, cyclone-track and wind-power targets. Probabilistic information can support risk decisions, but retrospective skill does not guarantee accurate prediction of every extreme event or operational outcome.
A staggered rollout covering 5,172 support agents found about 15% more issues resolved per hour with AI assistance. Less experienced and lower-skilled workers benefited most; experienced staff saw smaller speed gains and some quality declines. Evidence from one firm does not establish economy-wide employment or wage effects.
In controlled debates involving 900 participants, personalized GPT-4 was more persuasive than the human baseline. A 2026 correction clarifies that its direct advantage over non-personalized GPT-4 was not statistically significant. The study therefore cannot isolate a personalization benefit, nor establish long-term persuasion effects on real platforms.
The ILO combines task data, worker assessments, expert input and model scoring to refine occupational exposure estimates. Around one in four workers holds a job with some generative-AI exposure, with clerical work highly exposed. The index maps potentially affected tasks rather than forecasting layoffs; reorganization and full replacement are different outcomes.
This synthesis examines general-purpose AI capabilities, risks and mitigation, separating documented harms from uncertain future risks. It highlights an evidence dilemma: premature interventions may fail, while waiting for certainty can leave serious risks unmanaged. A research synthesis does not make every discussed scenario an inevitable outcome.
In a Swedish randomized trial involving roughly 106,000 women, AI-supported screening improved sensitivity, maintained similar specificity and reduced reading workload. Follow-up interval-cancer rates met non-inferiority criteria, but the between-group reduction was not statistically significant. This evaluates a specific workflow, not mortality reduction or fully autonomous screening.
In a high-school mathematics experiment, both a standard GPT interface and a tutoring variant improved assisted practice. After tool removal, the standard-interface group performed worse independently, while tutoring safeguards largely mitigated this loss. Design affected short-term learning in this setting, not necessarily long-term outcomes across all subjects.