The first entry is always the pin's source. Overall confidence is a weighted average of how firmly each reference supports the start and end times used above; a reference counts half as much for every 180 days older than the newest.
OpenAI's own model release notes date this entry September 12, 2024
AIandBlockchain's episode "AI's Quantum Leap Unpacking OpenAI's[1] O1 Models" (2024-10-07) says: "a bit unsettling. Join us as we break down the key elements of OpenAI's O1 models, using excerpts from their recently released system card ( September 12th ). We’ll discuss the unique strengths of the O1 Preview and O1 Mini models, highlighting their enhanced ability to handle complex coding tasks, while also"
Archived copy of OpenAI's[1] o1-preview announcement, which says "GPT-4o correctly solved only 13% of problems, while the reasoning model scored 83%", o1-preview scored 84 on a jailbreaking test, and o1-mini "is 80% cheaper than o1-preview".
Something missing or wrong? Say it in your own words: a link that backs this pin up, a different start or end date and why, or a fact it lacks or gets wrong. The AI checks it against this pin's sources, searches for better ones, and adds any page that backs you up. The pin's own sources still count most.