GitHub is having trouble counting things

(chuckgreenman.com)

72 points | by chuckgreenman 5 hours ago

24 comments

  • nxc18 3 hours ago
    GitHub, including Enterprise (so not just Azure’s fault), has become unreliable, low quality software.

    I say unreliable because I cannot rely on it to accurately do what it claims to do. I cannot trust any number. I cannot trust that actions I invoke will actually happen. I cannot trust that taking action will not have mysterious side effects (e.g. re-opening an issue my boss’s boss closed in 2017 without explanation). I think I trust the underlying git infra but at this point I’m not sure why I do.

    I say low quality because it is extremely slow at random moments, gets into broken inconsistent UI states, has baffling UX choices that make it harder to navigate than it should be (particularly with new UIs), and overall just doesn’t have the fit and finish we should expect from software in 1996 - or 2026, or literally any time in between.

    • plasma_beam 3 hours ago
      For months now I've had scheduled GH Actions workflows that don't run on time, sometimes not at all. Like we're talking 4-8 hours later than the cron statement. Yesterday I finally gave up and just moved the trigger to AWS EventBridge Scheduler/Lambda. Maddening, but also we're talking code repos that are $0 for me. Not production level. I can't imagine people running Production pipelines on github.com.
      • itsmeduncan 49 minutes ago
        My scheduled for 8am PT GH actions workflows don't run until anywhere between 6-8 hours later. I gave up, and run them locally now. It is just to create a tag, and push to a specific place. A super simple problem that was solved for so long simply with GH Actions that just doesn't work anymore.
      • paradox460 2 hours ago
        At a previous gig we had a simple "reset staging" task on gha, that ran once a week on Tuesday

        It ran fine for years, then around November just stopped and never ran again. Tried moving it around for which time it would trigger, and it just never started again. Opened a ticket, and got nowhere

      • DetroitThrow 2 hours ago
        I've switched off to runners and event scheduling inside AWS, and am working on moving my company off of GitHub. I know engineers inside GitHub and it doesn't sound like the talk they've been putting out about reliability is actually being addressed with many resources internally; the best people in the company are chasing AI product features (which are apparently quite profitable).
    • shmoil 1 hour ago
      >> GitHub has become unreliable, low quality software.

      They are owned by Microsoft, what did you expect exactly?

  • zX41ZdbW 3 hours ago
    Some of the oldest PRs could no longer be found.

    For example, when I decided to continue my eight-year-old PR https://github.com/ClickHouse/ClickHouse/pull/104948, I couldn't see it - neither in search nor while navigating through pages.

    Direct links still work.

  • datadrivenangel 3 hours ago
    Dirty reads is a classic symptom of scaling poorly. Eventually consistent database writes are fine, but it's a bad look if a user takes an action and then doesn't see the results of their action.
    • wren6991 3 hours ago
      Similar to how every description of relaxed consistency for CPUs starts out by pointing out your own reads always observe your own writes in program order, as if to say, "don't worry, we're not insane."
      • BoingBoomTschak 34 minutes ago
        Or funnier version: "don't worry, we're not Alpha".
  • Droobfest 2 hours ago
    I was having the problem yesterday that the GitHub API sometimes simply doesn't list an in progress workflow run in the output when I specifically filter for it:

    "api.github.com/repos/{owner}/{repo}/actions/runs?status=in_progress" -> outputs 3 workflow runs

    5 seconds later -> outputs 2 workflow runs

    5 seconds later -> outputs same 3 workflow runs again

    Apparently they make no guarantees of any type of consistency or continuity in output, which has the side effect of also making it completely useless.

  • huurtehoog 3 hours ago
    The fact that there's so much indirection between what happens in the registers and memory in the machine, the underlying reality being modeled by software, and the information displayed semantically to the users, is a travesty.

    Software could be so much simpler and more reliable. There's so much bloat that is necessary to solve problems created by bloat. I hope we find our way out of this mess sooner rather than later. I'd hate for entire generations to suffer the current state of software whereas the theoretical understanding necessary to make things better was produced very early in the history of programmable computers.

  • hanspagel 4 hours ago
    I think they struggle with two things

    1) counting

    • eugenekay 4 hours ago
      2) cache validation

      3) off-by-one errors

      • rhdunn 2 hours ago
        5) multithreading

        4) database synchronization

        • zbentley 34 minutes ago
          Ideally written as:

          5) multit

          6) database synchreadinghronization

      • kibwen 3 hours ago
        According to the documentation there are four things:

        4) naming things

        5) keeping the documentation up-to-date

      • Gander5739 3 hours ago
        There are solutions: https://xkcd.com/3062/
      • lbanchio 3 hours ago
        4) naming things
    • stateoff 4 hours ago
      "Since we forced our engineers to agentic workflows our efficiency doubled. Just look at the numbers!"
  • stabbles 2 hours ago
    The reason is that their primary source of truth (traditional database) and their search index (elasticsearch) are out of sync.

    The issues/pulls pages used to be showing the data from the primary data source, and then they changed it so everything is search, including the basic props is:pr and state:open.

    • ako 2 hours ago
      Don't worry, it will be consistent eventually.
    • oxidant 2 hours ago
      Experienced this using the API. Got told it's a "wontfix"
  • drdexebtjl 2 hours ago
    These days I was surprised to see, on my personal computer, information from PRs in private repos from my work’s org on GitHub. These are usually secured behind SSO flows which I don’t/can’t do on personal devices. So now I also don’t trust them to keep our code private anymore either. Nice.

    It used to be that you could have a single GitHub account and join multiple orgs, some of which have enterprise plans with GitHub. They completely fumbled this during the pressure to ship AI features. Some enterprise controls that are supposed to affect only my work at a specific GitHub org affects me account-wide. I’m part of quite a few orgs on GitHub for OSS projects, and I never consented for that one org’s policies to basically take over my account.

    I thought of going through the process of creating another account specifically for my work at this company, but GitHub is itself solving the problem in a very unique way: by being a place I don’t want to be in after I clock out.

  • Rooster61 3 hours ago
    > Really curious as to what’s going on over there, if you’ve got any insight let me know!

    Microsoft. That's the answer. Microsoft is happening over there.

  • cedws 3 hours ago
    There's tonnes of UI bugs. A more severe one I've experienced is that my approved PRs have been shown as still waiting on review, leading some of them to be delayed by weeks.
    • esafak 3 hours ago
      I'm partial to the one where the workflow completes but the its summary on the PR does not get updated, so it appears to be running.

      How many companies engender taste in bugs, eh??

  • bob1029 2 hours ago
    This appears to be a security problem in some contexts. I've been able to see issue counts for repositories that I have not been granted access to yet (the view with the invite accept button). I can't actually get to the issues but I can see how many there are.
  • thetnaingnyc 2 hours ago
    GitHub's downfall must be studied. At the same time as all of these reliability issues, they've been introducing a variety of small UI updates that don't make any meaningful updates either...
    • bakugo 2 hours ago
      That's just AI at work. It's much easier to ask Copilot to make some unnecessary UI change or add a new minor feature nobody needed, than it is to ask it to fix the fundamental stability and reliability issues that plague the software.

      And knowing how these sorts of large corporations are structured, I wouldn't be surprised if the people who spend a few hours prompting AI to add some new unnecessary feature are being rewarded more than the people breaking their backs trying to stop the site from collapsing under the weight of 100 million vibe coders.

  • stack_framer 2 hours ago
    The same counting problem exists with GitHub issues. I notice this every time I close an issue, and the count of open issues is still one higher than the actual number (until I reload the page).
  • apocalyptic0n3 3 hours ago
    Calling this "severe" feels overblown. It's a UI mismatch likely caused by using multiple caches that are out of sync. I feel like every engineer on this site has encountered this exact bug at one point or another. It's an easy one to run into, especially in bigger orgs.
    • layer8 3 hours ago
      The second screenshot with “1 3 Next” seems pretty hard to justify.

      Taking care that such incongruent representations can’t happen even in the face of inconsistent backend caches is part of proper UI state management.

    • thibaut_barrere 3 hours ago
      Yes but in the case of GH, it also happens in cases where the data is old: not just a temporary drift.
  • meerita 2 hours ago
    I know this pain because, right now, I have zero PRs opens, yet the PR tab is always saying I have 1. It's plainly stupid.
    • busymom0 2 hours ago
      The chat notifications count badge on Reddit website has the same problem for me. Tells me I have messages when I don't. Or won't tell me I have messages when I do.
  • nubinetwork 3 hours ago
    My phone consistently says that I have one more update than they'll actually let me download... what that one phantom app is, I'll never know...
  • disko 3 hours ago
    I have got an insight alright. Microsoft took it over.
    • herbst 2 hours ago
      Soon most will have forgotten how essential and stable github was at some point in the past.
  • vanyauhalin 2 hours ago
    I have been seeing this for one-two years already. They cannot count how many packages I have.

    https://github.com/vanyauhalin?tab=packages

  • mococa 3 hours ago
    AI will solve this
  • booi 3 hours ago
    Press "Next"
  • mickael-kerjean 3 hours ago
    This got to be the extinguish phase from the EEE Microsoft playbook. Prior acquisition, Github was very liked and very focus, then Microsoft happened in their embrace phase, telling us EEE was something of the past. While trying to show good faith they extended the platform and this is now the very last phase
    • yodon 3 hours ago
      >EEE

      My god what an ancient trope.

      Presumably you realize most current Microsoft employees were not even alive when the EEE memo was written.

      Must we also talk about some decision Henry Ford made in 1906 every time the automaker is mentioned, or something the Gauls did in 200 BC every time England is mentioned?

      • pessimizer 3 hours ago
        You don't want to talk about it because either 1) you don't know it, or 2) you were downplaying it at the time, too.

        The only time we talk about Henry Ford, unless it's about his anti-Semitism, is about decisions he made in 1906. If we go by your arbitrary standard, why would we talk about Henry Ford at all when we could talk about this year's Tesla model?

        > See also Netscape v2.0 and 3.0, which were badass browsers back in 1996, but are similarly irrelevant today.

        Why are you talking about Netscape? What does that have to do with anything happening right now?

    • chuckgreenman 3 hours ago
      Maybe, but does Team Foundation Server even exist anymore? What is the MS branded service that degrading GitHub is supposed to drive people to?
      • isolay 3 hours ago
        Copilot, of course. Helpfully, every Microslop product is named Copilot Something now. Take your pick.
    • herpdyderp 3 hours ago
      I don't believe this is true of the past, but it'd be quite the twist if the extinguish phase was really just pure incompetence this whole time. It certainly seems that way right now with GitHub.
  • whalesalad 3 hours ago
    rails + russian doll style caching + large engineering team + azure as a platform = recipe for this exact problem.
  • crote 3 hours ago
    [dead]