4 comments

  • dada216 13 minutes ago
    We built a complete production-grade inference service from scratch on a cluster of more than 100,000 Chinese-made AI accelerators. All production inference for GLM-5.3-Flash runs on this system.
  • tefkah 3 minutes ago
    > Today, GLM-5.3 has become an indispensable daily coding partner for everyone on the team, and it is moving steadily toward replacing us. If this trend continues, given enough compute and enough time, its endpoint is a system that can design and train its own successor entirely autonomously. This is known as Recursive Self-Improvement, or RSI.

    Statements dreamed up by the utterly deranged.

  • bbor 3 minutes ago
    Well, other than the infrastructure they got from illegally routing millions of paying customers' requests through Anthropic's Opus 4.8 in a distillation attack...
  • embedding-shape 10 minutes ago
    I was gonna ask how people found their coding plans, and realize, have they massively ramped up the prices? Seems the middle plan is ~$80/month now, didn't that used to be like $20/month? Cheapest plan is ~$20/month currently.

    They must have hit really hard scaling limits if the prices were hiked so much so quickly.

    • broodbucket 7 minutes ago
      Yeah it went from a great deal to unviable compared to other providers imo. They really need to find a healthy middle ground