![]() |
|
|
|
#1
|
||||
|
the machine god weighs in
Quote:
| |||
|
#2
|
|||
|
Now that the (largely) freebie tokens will dry up because billionaires want they money back those of you that ask A.I agents stupid questions like Summarize the posts of forum user <insert user> and give me bullet points as to why they are a boomer for free will defo not wanna pay for that shit.
A.I agents need to make money for their digital pimps. Start paying. Medical imaging manipulation/interogation? Good usecase. Vibe coding with guided prompts by knowledgeable programers? As much as i hate to admit it. Good usecase. Putting a suit and tie on ya cat photo? Go for it, but be prepared to pay for it with little to no returns. Welcome to A.I. | ||
|
#3
|
||||||
|
that graph now that ive actually look at it is just about time on task and stops at 16 hours dues to unreliable test suite, whatever the fuck that means, write a better test, either way its outdated as fuck Claude runs for days now
Quote:
the thing that actually matters is capability and the plateau is way more pronounced in those charts in the models released in the last year, the big gains are in reducing the time to complete a task successfully, some giant codebase can take chatgpt 5.5 3 days to work on and Mythos supposedly did the work correctly in like 10 hours or something [You must be logged in to view images. Log in or Register.] we've hit that ceiling, the is AGI even possible easily by just making a 1 trillion parameter model type idea and the answer is no and charts like this are pointless now because of diminishing returns of just training larger and larger parameter single models Qwen opensource is at like 350b parameters but that's probably just adding up all the separate MoE models Quote:
Quote:
so not only is AGI/ASI not going to happen, all these companies are going to go bankrupt causing a deep recession because their business plan we started this journey with doesn't make any sense anymore open source models wins on both ends of the spectrum, locally run open source wins for a non trivial chunk of the consumer/enthusiast market and enterprise coding just got something that costs 1/5th of a Claude or ChatGPT seat dropped in their lap thanks to China | |||||
|
#4
|
||||
|
Quote:
| |||
|
#7
|
||||
|
Quote:
I'm not saying these researchers are everyday users. These are the exact senior scientists at Google DeepMind who build and evaluate Gemini. The chart isn't supposed to show how a casual user prompts a model; it's a technical evaluation from the actual creators of the AI showing how the underlying architecture behaves. They absolutely 'matter' because they build the tech. | |||
|
#8
|
||||
|
My beef is with the METR graph which isn't mentioned once in the paper nor are they, its a sensational Berkeley nonprofit writing disingenuous misleading tests to show the outcome they want, working backwards from ai in scifi scary so we should stop just like they work backward from cows fart too much so you shouldn't be allowed to have a hamburger
[You must be logged in to view images. Log in or Register.] If you do click one of the dots it does list way past 16hours shown but the methodology of the test itself and the chart are both misleading to show scary exponential growth, nothing changed with how the models themselves are fundementally built, it's stuff around the model that is improving, the harness & MoE. You can stick a earlier model in a harness and let it run for days also but the chart doesn't show that they cap gpt5 at 6 hours and previous models are 12-30 minutes, if you run the same opus they have ranked as they do without a harness it scores will be completely different Quote:
https://summify.io/discover/is-ai-ab...-s-not-5GezB1/ one click bait YouTuber to counter another And the Google fanfic thought expirement about the timeline of one sci fi concept progressing into another sci fi theoretical concept itself I have no issue with | |||
|
#9
|
|||||
|
I posted a document from the lead developers on Gemini, not a youtuber.
Quote:
Quote:
This is the part I liked anyway. | ||||
|
#10
|
||||||||
|
Quote:
Quote:
Quote:
Quote:
Quote:
this is the same company that had a mustard tiger named Blake Lemoine who thought a now obsolete chatbot from years ago had a fucking soul because of his religious views, smart people can be retarded and have views on super ai gonna kill us all or turn the entire universe into paperclips, one dipshit taken seriously until recently was even afraid of the concept of a ASI in the future with time traveling capabilities that would torture him for eternity for not working on ai, these are just thought experiments same as AGI/ASI | |||||||
![]() |
|
|