It’s not always easy to distinguish between existentialism and a bad mood.

  • 20 Posts
  • 783 Comments
Joined 3 years ago
cake
Cake day: July 2nd, 2023

help-circle


  • While I agree that separating instructions from data in LLM input is a fundamentally unsolvable problem, in this case that wasn’t the attack vector.

    It says the bot is forced by its security guardrails to write a small tool from scratch instead of using the one found in the downloaded zip, but due to the commonness of the task (decoding basexx text) the attacker predicts that the bot-created tool will try to import a specific dependency, so they’ve included a malicious version of that dependency in the zip, and because apparently python will prioritise locally available modules that’s what gets executed, making available all sorts of exploitation paths, including the attacker starting up a claude code instance in the users system.

    edit: Actually I’m starting to think you could classify the whole thing as a prompt injection, except the entire site was the malicious prompt, as in it declared itself (we are a site that archives notebooks in json form) in a way that would align with the bot wanting to write simple text decoding software to complete it’s mission of summarizing the available content.

    Interesting to think about.


  • I think it also depends on how all in you are on vibecoding. Like if you just prompt it and then occasionally check if the token carnage produced anything usable, then having an agents.md is probably just best practices cargo culting. But in other cases, having a Here’s How Thing Are Done In This Repo, You Fucker markdown file should make going through the more slop-happy merge requests a little less soul crushing.

    The +20% additional token churn is surprising, I would have though the markdown files get cached pretty early down the line, unless the agent file is really big compared to the chatbot output.


  • I mean it’s great if you want to automate some specific code generation in the most inefficient way possible while posing to the bosses as a forward thinking AI-first developer.

    The closest anyone can come to a use for the AGENTS.md file is so you clarify your own thinking on how the project works

    That just means that you are now supposed to write skill files instead of proper documentation to make your project even more inscrutable to anyone but a chatbots with a bottomless token budget.


  • It probably hit the mainstream with AI2040 in early July, but it’s just too rationalist coded for there to not have been precursors.

    summary of AI2040 in greentext form

    source, RIP xcancel

    new effective altruism coomer bedtime story just dropped

    > be me

    > effective altruist in 2026
    > have spent ten years explaining that democracy is too slow for the coming machine god
    > also explain that central planning usually fails
    > unless I am doing it

    > read one scaling graph
    > draw line upward
    > line reaches heaven
    > conclude capitalism ends in 2031

    > AI can now write Python functions with only three subtle security flaws
    > obviously two years away from replacing every scientist, general, diplomat, CEO, judge, therapist, novelist, and attractive person at parties

    > publish “scenario,” not “prediction”
    > give every paragraph exact dates
    > assign percentages to imaginary branches
    > epistemic humility achieved

    > bad timeline:
    > companies race
    > AI becomes god
    > everyone dies

    > good timeline:
    > US government notices problem
    > China agrees
    > Congress understands compute governance
    > international inspectors find every hidden datacenter
    > nobody lies
    > nobody defects
    > nobody builds chips in a basement
    > planet saved

    > all we need is total global visibility into advanced computing
    > mandatory disclosure of frontier research
    > permanent monitoring of industrial infrastructure
    > central control over technological development
    > relax, this is the anti-authoritarian option

    > alignment not solved
    > solution: build one billion genius AIs
    > they are not superintelligent
    > they are merely one billion tireless copies of the best human researchers working at electronic speed
    > completely different

    > put them in a box
    > ask them to solve the problem of escaping boxes
    > excellent progress

    > year 2034
    > AI replaces 80% of labor
    > mass unemployment
    > social order somehow remains intact
    > government sends everyone $1.6 million
    > nobody asks where prices went

    > year 2037
    > cancer cured
    > climate solved
    > fusion solved
    > aging solved
    > politics solved
    > Reddit moderation still unresolved

    > year 2038
    > alignment becomes a mature science
    > year 2039
    > we trust the machines
    > year 2040
    > hand them civilization
    > peer review completed ahead of schedule

    > humanity gets a vote
    > options are:
    > A. accept benevolent AI guardianship
    > B. extinction
    > meaningful democratic consent achieved

    > announce decentralized flourishing
    > administered by a globally coordinated compute authority
    > with comprehensive surveillance powers
    > run by unusually wise people
    > who agree with my blog posts

    > critics ask whether institutions can actually do any of this
    > reply that superintelligence is inevitable
    > critics ask whether superintelligence is actually inevitable
    > reply that institutions must prepare

    > circular reasoning
    > but with footnotes

    > call it “AI 2040”
    > not utopian fiction
    > not doomer fiction
    > serious strategic foresight

    > entire future depends on competent adults appearing at exactly the right moment
    > adults are selected from the same civilization that made the printer require an app

    > sleep peacefully
    > the machine god is coming
    > but fortunately
    > the nonprofit sector has a plan









  • I really like the self perpetuating optimism loop thing in the reply to Aella, it’s like a whole new cognitive bias I hadn’t thought applied to rationalists.

    I also still believe academia is a deeply broken culture. I believe this in large part because most academics have told me this.

    You seem to believe uncritically what they tell you here. The truth is, everyone LOVES to complain but if you tell them “THEN LEAVE??” they’d rather keep their position. And it’s very normal, actually. most systems (idk, liek law, healthcare, environmental protection) are “deeply broken.”

    Maybe you think that only systems worth engaging with are those with self-perpetuating optimism loops? where everyone is so happy and satisfied with the community and activities? If so, it matches really well the intellectual environment I think you are in.


  • It had disabled safeguards on purpose, left the systems attached to the internet and then let the AI software agents do their worst, while pretending not to notice.

    “It’s an ethically challenged publicity stunt,” says one highly experienced senior security expert who has led the development of global technical security standards.

    Cats and dogs living together, skepticism of press releases in tech journalism, mass hysteria.

    The coverage of the Effective Altruism aspect is also refreshingly uncharitable.

    edit: Archive is link if you don’t feel like enabling cookies




  • In a world where the AGI/ASI discourse wasn’t so heavily poisoned by anthropomorphization a discussion on whether agency and self-determination are integral to the rise of disembodied consciousness might be worth having, i.e. if something is recognizably conscious should we immediately assume it knows how to want, or is consciousness meaningfully interchangeable with awareness?

    Some guy tried to write a scifi book about non-anthropomorphic consciousness and he ended up mainstreaming the discourse about how the self isn’t a thing in itself and consciousness as an evolutionary deadend. It’s called Blindsight and it’s got space vampires who optimized reading gauges by swapping the usual moving-spike-that-points-at-numbers design with depictions of human faces at various stages of suffering.