• 0 Posts
  • 25 Comments
Joined 1 year ago
cake
Cake day: June 28th, 2025

help-circle


  • Mastodon has some weird etiquette I don’t quite understand. Maybe it’s because I’ve never really used “microblogging” or much social media outside old forums. Most people seem to use it for self-promotion. If you respectfully disagree on some part of someone’s post or just try to be helpful and add more context, sometimes the OP gets upset (I’ve heard the term “reply-guy” mentioned on there before, and also seen “experts” pulling the “do you know who you are talking to” card when others try to add context). There is little discussion and interaction from what I’ve seen; mostly just posts without meaningful replies.

    Politically, Lemmy seems to lean left-ish. Mastodon leans liberal, but there’s more variety than Lemmy (I guess since interaction is low). I haven’t used Bsky, but I saw when Stephen King announced he voted for Platner he got dogpiled (dunno if that means Bsky leans establishment Dem that would prefer zionist and corporate Janet Mills or what).






  • sobchak@programming.devto196@lemmy.blahaj.zonerule
    link
    fedilink
    English
    arrow-up
    2
    ·
    3 months ago

    I don’t. But, I don’t think posts about oral sex are typically the top post on the equivalent of “All” on other social media, as this post was. I could be wrong. I routinely see sexual topics on All here (sorted by Active). Nothing wrong with it, just making an observation.




  • 14 H100 GPUs per user that they serve

    Not per user, but probably decent rough estimate to that per vibecoding dev that is continually running agents 8+ hours/day. Some people’s “workflows” involve running multiple parallel agents sometimes or even a significant portion of the time (using the git worktree feature), so I think that’s probably a decent rough estimate. I imagine the limit would be serving 10 of these types of “devs.” Of course, there’s batching and stuff that can be done, but I think it still slows everybody else down near linearly. H100s aren’t the only accelerators used for inference; I just chose it as an example. Google has ~5 million H100 equivalent accelerators, Microsoft has 3.5 million, and Amazon has 2.5 million (https://www.networkworld.com/article/4156949/google-owns-the-most-ai-compute-and-it-built-it-its-way.html).


  • sobchak@programming.devtoPeople Twitter@sh.itjust.worksManagers
    link
    fedilink
    arrow-up
    3
    arrow-down
    4
    ·
    edit-2
    4 months ago

    Probably more expensive than the subsidized costs. Hmm…

    H100 GPUs cost $25k, and have 80GB of RAM. Kimi k2.6 has 1.1T parameters. Assuming 8 bit quantization, would need 14 GPUs to run a single agent at a time (I’m not sure the cloud models use quantization; it could be double). So, $350k per vibecoding dev on GPUs alone. Life expectancy is ~4 years, so ~90k/year amortized. This is ignoring the significant electrical/HVAC cost of handling 10KW of electricity and heat per vibecoding dev (and tons of other costs as well).