[-] brianpeiris@lemmy.ca 8 points 5 days ago

Bless uncle Bernie, but the US is in no state to even consider this. It has installed capitalism as its king.

-12

Summary:

  • GPT-6 Astra scores 62.7% for $26K on ARC-AGI-3 Semi-Private with our Standard harness, and 99.9% for $19K with a Provider Adapter harness.
  • GPT-6 Astra surpasses the human baseline in action efficiency on ARC-AGI-3. It used fewer actions than the median tested human on 96% of levels.
  • A key behavior observed in GPT-6 Astra was its ability to turn unfamiliar environments into compact symbolic world models. It represented game mechanics as logical rules and developed its own domain-specific language shorthand to track state and plan actions.

For a cost comparison, during our controlled testing, human participants were paid $115 per 90-minute session, plus $5 per game completed. Participants attempted approximately nine games per session, roughly $12.78 per attempted game before bonuses.
Most of this fee pays for the participant’s time and willingness to take the test, rather than the energy their brain uses (a closer proxy to compare with AI). If we look at only the brain’s energy, and price it as electricity, the estimate drops to about 0.6 cents per session, or 0.067 cents per game attempted.1

Astra’s results are also a major milestone worth celebrating. From our perspective, Astra represents a noticeable step-function change in frontier model capabilities.

When we launched ARC-AGI-3, we made it clear that saturating the benchmark would not represent “proof of achieving AGI.” Therefore, while we believe Astra represents meaningful progress towards generalization, we are not claiming that it is AGI.

-26
submitted 2 weeks ago* (last edited 2 weeks ago) by brianpeiris@lemmy.ca to c/technology@lemmy.world

Summary:

  • GPT-6 Astra scores 62.7% for $26K on ARC-AGI-3 Semi-Private with our Standard harness, and 99.9% for $19K with a Provider Adapter harness.
  • GPT-6 Astra surpasses the human baseline in action efficiency on ARC-AGI-3. It used fewer actions than the median tested human on 96% of levels.
  • A key behavior observed in GPT-6 Astra was its ability to turn unfamiliar environments into compact symbolic world models. It represented game mechanics as logical rules and developed its own domain-specific language shorthand to track state and plan actions.

For a cost comparison, during our controlled testing, human participants were paid $115 per 90-minute session, plus $5 per game completed. Participants attempted approximately nine games per session, roughly $12.78 per attempted game before bonuses.
Most of this fee pays for the participant’s time and willingness to take the test, rather than the energy their brain uses (a closer proxy to compare with AI). If we look at only the brain’s energy, and price it as electricity, the estimate drops to about 0.6 cents per session, or 0.067 cents per game attempted.1

Astra’s results are also a major milestone worth celebrating. From our perspective, Astra represents a noticeable step-function change in frontier model capabilities.

When we launched ARC-AGI-3, we made it clear that saturating the benchmark would not represent “proof of achieving AGI.” Therefore, while we believe Astra represents meaningful progress towards generalization, we are not claiming that it is AGI.

30
submitted 2 weeks ago by brianpeiris@lemmy.ca to c/canada@lemmy.ca

Lawyers representing victims and survivors of the Tumbler Ridge school shooting say they are filing 30 new lawsuits against OpenAI in U.S. federal court in California.

The lawsuits accuse OpenAI of "choosing profit over the lives of the children of Tumbler Ridge" as it prepares to become a publicly traded company.

[Jay Edelson] said OpenAI has so far refused calls to make the logs public, a choice he believes indicates their contents may not work in the tech company's favour.

32
submitted 2 weeks ago by brianpeiris@lemmy.ca to c/fuck_ai@lemmy.world

Lawyers representing victims and survivors of the Tumbler Ridge school shooting say they are filing 30 new lawsuits against OpenAI in U.S. federal court in California.

The lawsuits accuse OpenAI of "choosing profit over the lives of the children of Tumbler Ridge" as it prepares to become a publicly traded company.

[Jay Edelson] said OpenAI has so far refused calls to make the logs public, a choice he believes indicates their contents may not work in the tech company's favour.

36
submitted 2 weeks ago by brianpeiris@lemmy.ca to c/fuck_ai@lemmy.world

University’s move is part of trend that critics within academia say will lead to further staff cuts and the loss of ‘everything that makes the job worth doing’

One student, who asked to remain anonymous, said they felt “short-changed” after issues with their degree structure forced them to enrol in the online option.

“Instead of a group discussion or a tutor teaching you, it’s a class activity done with a robot,” the student said. “I feel like I’m teaching this robot how to take jobs away from my teachers.”

33
submitted 1 month ago by brianpeiris@lemmy.ca to c/canada@lemmy.ca

The StatsCan report: Perceptions of gender-based violence and gender equality, identity and expression in Canada, 2025

Excerpt:

A new Statistics Canada report says Canada is becoming less accepting of gender diversity.

In 2025, 77 per cent of women and 70 per cent of men surveyed said they agreed or strongly agreed that people should be free to express their gender however they choose.

That’s a drop from 2018, when 85 per cent of women and 78 per cent of men agreed.

In Saskatchewan, 69 per cent of women and 60 per cent of men surveyed last year said they agreed, making it the province with the lowest acceptance of gender diversity in the country.

Women in Saskatchewan specifically were down 14 percentage points from 2018, when 83 per cent agreed people should be free to express their gender however they choose.

Since 2018, Saskatchewan and Alberta's provincial governments have passed legislation targeting trans and gender diverse youth, such as Saskatchewan’s Parents’ Bill of Rights, enacted in 2023.

100
submitted 1 month ago by brianpeiris@lemmy.ca to c/canada@lemmy.ca
23
submitted 1 month ago* (last edited 1 month ago) by brianpeiris@lemmy.ca to c/technology@lemmy.world

The top three solutions come from independent researchers. The best solution was built by a group of PhDs and professors, who released a corresponding paper. They all make use of some form of world-model.

I've generally been a skeptic, and I still am, but this news surprised me because I expected ARC-AGI-3 to remain difficult for a long while.

Note that the scores are self-reported and need to be independently verified. The solutions have not been tested against the larger private test set.

Primer on ARC-AGI-3:

ARC-AGI-3 is an interactive reasoning benchmark which challenges AI agents to explore novel environments, acquire goals on the fly, build adaptable world models, and learn continuously.

A 100% score means AI agents can beat every game as efficiently as humans.

Instead of solving static puzzles, agents must learn from experience inside each environment—perceiving what matters, selecting actions, and adapting their strategy without relying on natural-language instructions.

65
submitted 1 month ago* (last edited 1 month ago) by brianpeiris@lemmy.ca to c/canada@lemmy.ca
35
submitted 2 months ago by brianpeiris@lemmy.ca to c/canada@lemmy.ca

Amnesty International Canada is opposing the project by Meta, which is anticipated to take up an area larger than Vancouver’s Stanley Park in Sturgeon County.

“This is a highly unregulated industry that operates on a surveillance-based business model,” Tara Scurr, Amnesty International Canada’s Corporate Accountability and Climate Justice Campaigner.

161
submitted 2 months ago by brianpeiris@lemmy.ca to c/fuck_ai@lemmy.world

Some families in Georgia are being forced to sell their homes or face government seizures to make way for powerlines. 70-80% of that power is needed for AI data centers.

[-] brianpeiris@lemmy.ca 73 points 2 months ago

If the Zig community carves out a territory of principled engineering like this, I may adopt it as my primary language and make a career out of it. Finally an island of sanity in a sea of slop.

38
submitted 2 months ago by brianpeiris@lemmy.ca to c/canada@lemmy.ca

In the past year, Canada's immigration rate has experienced a dramatic reversal. We explain how it happened.

Also on Nebula: https://nebula.tv/videos/tldrnewsglobal-canadas-insane-immigration-uturn-explained

[-] brianpeiris@lemmy.ca 35 points 3 months ago

Bernie has good intentions, but he was AI-pilled by Geoffrey Hinton, who ironically also has good intentions. However, they are both out of touch with reality.

[-] brianpeiris@lemmy.ca 106 points 3 months ago

I like Ed, but not a fan of this style of teasing. Reminds me of conspiracy theory communities. We'll see what he has I guess.

[-] brianpeiris@lemmy.ca 72 points 4 months ago

The Luddites didn’t hate machines. They were gifted artisans resisting a capitalist takeover of the production process that would irreparably harm their communities, weaken their collective bargaining power, and reduce skilled workers to replaceable drones as mechanized as the machines themselves. Their struggle has been tragically warped into a caricature when it is more relevant than ever.

https://www.currentaffairs.org/news/2021/06/the-luddites-were-right

[-] brianpeiris@lemmy.ca 35 points 5 months ago

Buddhist Copilot builds apps with sublime coding standards, and on the last iteration it runs rm -rf * .git before it recites a koan on impermanence.

[-] brianpeiris@lemmy.ca 36 points 5 months ago

Looks like the maintainer burned out. Maybe give them some time to recover.

https://github.com/nvim-treesitter/nvim-treesitter/discussions/8627

[-] brianpeiris@lemmy.ca 57 points 7 months ago

Good reminder to donate to web.archive.org

[-] brianpeiris@lemmy.ca 44 points 7 months ago

Not a fan of Poilievre by any means, but I'm glad we don't live in a world where he immediately takes the anti-trans attack angle. I won't be surprised if he does in a few days or weeks, but I'll take what I can get.

[-] brianpeiris@lemmy.ca 42 points 10 months ago* (last edited 10 months ago)

Yes, probably, but you know what, even if the DEI was performative, it had a real positive impact for tens of thousands of employees and the culture set by the media empire they control, and now we don't even have that.

view more: next ›

brianpeiris

0 post score
0 comment score
joined 3 years ago