For OpenAI, o1 represents a step toward its broader goal of human-like artificial intelligence. More practically, it does a better job at writing code and solving multistep problems than previous models. But it’s also more expensive and slower to use than GPT-4o. OpenAI is calling this release of o1 a “preview” to emphasize how nascent it is.

The training behind o1 is fundamentally different from its predecessors, OpenAI’s research lead, Jerry Tworek, tells me, though the company is being vague about the exact details. He says o1 “has been trained using a completely new optimization algorithm and a new training dataset specifically tailored for it.”

OpenAI taught previous GPT models to mimic patterns from its training data. With o1, it trained the model to solve problems on its own using a technique known as reinforcement learning, which teaches the system through rewards and penalties. It then uses a “chain of thought” to process queries, similarly to how humans process problems by going through them step-by-step.

At the same time, o1 is not as capable as GPT-4o in a lot of areas. It doesn’t do as well on factual knowledge about the world. It also doesn’t have the ability to browse the web or process files and images. Still, the company believes it represents a brand-new class of capabilities. It was named o1 to indicate “resetting the counter back to 1.”

I think this is the most important part (emphasis mine):

As a result of this new training methodology, OpenAI says the model should be more accurate. “We have noticed that this model hallucinates less,” Tworek says. But the problem still persists. “We can’t say we solved hallucinations.”

  • Voroxpete
    link
    fedilink
    English
    arrow-up
    4
    arrow-down
    3
    ·
    2 months ago

    More and more advanced tools for automation are an important part of creating a post-scarcity future. If we can combine that with tearing down our current economic system - which inherently requires and thus has to manufacture scarcity - we can uplift our species in ways we can currently only imagine.

    But this ain’t it bud. If I ask you for water and you hand me a glass of warm piss, I’m not “against drinking water” for refusing to gulp it down.

    This isn’t AI. It isn’t - meaningfully and usefully - any form of automation at all. A bunch of conmen slapped the letters “AI” on the side of their bottle of piss and you’re drinking it down like it’s grandma’s peach tea.

    The people calling out the fundamental flaws with these products aren’t doing so because we hate the entire concept of automation, any more than someone exposing a snake-oil salesman hates medicine. What we hate is being lied to. The current state of this technology is bullshit and hype. It is not fit for human consumption (other than recreationally) and the money being pumped into it could be put to far better uses. OpenAI may have lofty goals, but they have utterly failed at achieving them, and right now any true desire to create AGI has been totally subsumed by the need to keep pumping out slightly better looking versions of the same polished turd in order to convince investors to keep paying for their staggeringly high hosting costs.