GitCode, a git-hosting website operated Chongqing Open-Source Co-Creation Technology Co Ltd and with technical support from CSDN and Huawei Cloud.

It is being reported that many users’ repository are being cloned and re-hosted on GitCode without explicit authorization.

There is also a thread on Ycombinator (archived link)

  • Grimy@lemmy.world
    link
    fedilink
    English
    arrow-up
    6
    arrow-down
    3
    ·
    7 days ago

    They should definitely respect the licenses, that being said, Microsoft owns GitHub and can be a bit quick in what they ban. It also means they are beholden to US laws, which could turn anti FOSS-AI in the near future.

    This is a smart move and I honestly hope more countries start doing it. It would probably lead to a better ecosystem.

    • TheGalacticVoid@lemm.ee
      link
      fedilink
      English
      arrow-up
      2
      ·
      7 days ago

      I think projects like this are good, but I really don’t want governments to create their own version of XYZ for the sake of creating clones of XYZ. I’m scared that all this will do is fragment an almost-universal collection of open-source projects into regional variants for no real reason.

  • raspberriesareyummy@lemmy.world
    link
    fedilink
    English
    arrow-up
    34
    arrow-down
    6
    ·
    6 days ago

    With the obligatory “fuck everyone who disregards open source licenses”, I am still slightly amused at this raising eyebrows while nearly no one is complaining about MS using github to train their copilot LLM, which will help circumvent licenses & copyrights by the bazillion.

    • Cosmicomical@lemmy.world
      cake
      link
      fedilink
      English
      arrow-up
      3
      ·
      6 days ago

      Came here to say this. As much as I don’t like china, there is really nothing to see (apart from the source, that’s for everybody to see).

      • mightyfoolish@lemmy.world
        link
        fedilink
        English
        arrow-up
        0
        ·
        edit-2
        6 days ago

        This could be illegal for git repos that do not have a open source license that allows mirroring or copying (BSD, Apache, Mit, GPL, etc.) Sometimes these repos are more “source available” and the source is only allowed to be read, not redistributed or modified. I would say that this is more of a matter for each individual copyright holder, not Microsoft.

        But ultimately I agree, this really isn’t as big of a deal as people are making.

        edit: changed some wording to be clearer

        • Maggoty@lemmy.world
          link
          fedilink
          English
          arrow-up
          0
          arrow-down
          2
          ·
          6 days ago

          China is a sovereign entity. I’m pretty sure they can decide foreign licensing laws don’t apply there.

          • mightyfoolish@lemmy.world
            link
            fedilink
            English
            arrow-up
            3
            ·
            6 days ago

            China is a soverign state and they should make their own laws. However, China has promised repeatably that they will take IP concerns more strictly (trade deal with Trump in 2020 is one example of this promise). It seems of this moment they still use the World Intellectual Property Organization for inspiration for their IP laws. At one point, China did not acknowledge IP rights at all. Being consistent is good for business; especially when it comes to international business.

            In 1980, China became a member of the World Intellectual Property Organization (WIPO). As of at least 2023, China’s view is that WIPO should be the primary international forum for IP rule-making. - Wikipedia

            • Maggoty@lemmy.world
              link
              fedilink
              English
              arrow-up
              2
              arrow-down
              1
              ·
              5 days ago

              China has never been consistent. Doing business there is all about relations with the CCP. This is a perfect example of how an authoritarian regime differs from a liberal regime. One is bound by it’s promises and rules and the other binds it’s rules to it’s needs.

    • Kusimulkku@lemm.ee
      link
      fedilink
      English
      arrow-up
      8
      ·
      5 days ago

      nearly no one is complaining about MS using github to train their copilot LLM

      What rock have you been living under??

    • JackbyDev@programming.dev
      link
      fedilink
      English
      arrow-up
      5
      ·
      6 days ago

      while nearly no one is complaining about MS using github to train their copilot LLM,

      Lots of people complained about that. I’ve only seen this single thread complaining about this.

    • kava@lemmy.world
      link
      fedilink
      English
      arrow-up
      0
      ·
      6 days ago

      If I look at a few implementations of an algorithm and then implement my own using those as inspiration, am I breaking copyright law and circumventing licenses?

      • sugar_in_your_tea@sh.itjust.works
        link
        fedilink
        English
        arrow-up
        1
        ·
        6 days ago

        That depends on how similar your resulting algorithm is to the sources you were “inspired” by. You’re probably fine if you’re not copying verbatim and your code just ends up looking similar because that’s how solutions are generally structured, but there absolutely are limits there.

        If you’re trying to rewrite something into another license, you’ll need to be a lot more careful.

        • kava@lemmy.world
          link
          fedilink
          English
          arrow-up
          0
          ·
          5 days ago

          What’s the limit? This needs to be absolutely explicit and easy to understand because this is what LLMs are doing. They take hundreds of thousands of similar algorithms and they create an amalgamation of it.

          When is it copying and when it is “inspiration”? What’s the line between learning and copying?

          • sugar_in_your_tea@sh.itjust.works
            link
            fedilink
            English
            arrow-up
            1
            ·
            5 days ago

            I disagree that it needs to be explicit. The current law is the fair use doctrine, which generally has more to do with the intended use than specific amounts of the text. The point is that humans should know where that limit is and when they’ve crossed it, with motive being a huge part of it.

            I think machines and algorithms should have to abide by a much narrower understanding of “fair use” because they don’t have motive or the ability to Intuit when they’ve crossed the line. So scraping copyrighted works to produce an LLM should probably generally be illegal, imo.

            That said, our current copyright system is busted and desperately needs reform. We should be limiting copyright to 14 years (as in the original copyright act of 1790), with an option to explicitly extend for another 14 years. That way LLMs can scrape comment published >28 years ago with no concerns, and most content produced >14 years (esp. forums and social media where copyright extension is incredibly unlikely). That would be reasonable IMO and sidestep most of the issues people have with LLMs.

            • kava@lemmy.world
              link
              fedilink
              English
              arrow-up
              0
              ·
              5 days ago

              First, this conversation has little to do with fair use. Fair use is when there is an acceptable reason to break copyright. For example when you are making a parody or critique or for education purposes.

              What we are talking about is the act of reading and/or learning and then using that information in order to synthesize new material. This is essentially the entire point of education. When someone goes to art school, they study many different artists and their techniques. They learn from these techniques as they merge them together in different ways to create novel art.

              Everybody recognizes this is perfectly OK and to assume otherwise is absurd. So what we are talking about is not fair use, but extracting data from copyrighted material and using it to create novel material.

              The distinction here is you claim when this process is automated, it should become illegal. Why?

              My opinion is if it’s legal for a human to do, it should be legal for a human to automate.

              • sugar_in_your_tea@sh.itjust.works
                link
                fedilink
                English
                arrow-up
                1
                ·
                5 days ago

                What we are talking about is the act of reading and/or learning and then using that information in order to synthesize new material.

                Sure, but that’s not what LLMs are doing. They’re breaking down works to reproduce portions of it in answers. Learning is about concepts, LLMs don’t understand concepts, they just compare inputs with training data to provide synthesized answers.

                The process a human goes through is distinctly different from the process current AI goes through. The process an AI goes through is closer to a journalist copy-pasting quotations into their article, which falls under fair use. The difference is that AI will synthesize quotations from multiple (many) sources, whereas a journalist will generally just do one at a time, but it’s still the same process.

    • ILikeBoobies@lemmy.ca
      link
      fedilink
      English
      arrow-up
      4
      arrow-down
      10
      ·
      6 days ago

      Are you just trying to make a bad pro-China argument or have you never been online before?

        • ILikeBoobies@lemmy.ca
          link
          fedilink
          English
          arrow-up
          3
          arrow-down
          3
          ·
          6 days ago

          “Why does no one say murder is bad unless China is murdering”

          Isn’t a good anti-murder argument

          • raspberriesareyummy@lemmy.world
            link
            fedilink
            English
            arrow-up
            6
            ·
            6 days ago

            “Why does no one say murder is bad unless China is murdering”

            I can not fathom how you absolutely nailed the essence of my comment, yet misunderstood it (and - arguably - your own example) so fundamentally.

            Let me try to help, once:

            “Why do most people not complain about murder when Microsoft is doing it, but when China is doing it, the very justified outrage can be heard?”

            • ILikeBoobies@lemmy.ca
              link
              fedilink
              English
              arrow-up
              2
              arrow-down
              1
              ·
              edit-2
              6 days ago

              I cannot fathom how you absolutely nailed the essence of my comment, yet misunderstood it (and - arguably - your own example) so fundamentally.

              People do criticize Microsoft for using open source data to train LLMs

              Hence the query about never been on the internet before

  • ikidd@lemmy.world
    link
    fedilink
    English
    arrow-up
    8
    arrow-down
    1
    ·
    7 days ago

    Shame they don’t have anything themselves that’s worth the trouble to copy back.

    • ZeroHora@lemmy.ml
      link
      fedilink
      English
      arrow-up
      3
      ·
      7 days ago

      Let’s dismiss all chinese contributors to open source projects with AI, javascript, PHP and so on.

      • ikidd@lemmy.world
        link
        fedilink
        English
        arrow-up
        6
        arrow-down
        1
        ·
        7 days ago

        That they got from the West when CATL bought out a bankrupt US company that had developed LFP to commercial viability.

        • IHeartBadCode@kbin.run
          link
          fedilink
          arrow-up
          7
          arrow-down
          2
          ·
          7 days ago

          I think the two of you are focusing on either end of this and not really seeing the bigger picture.

          China absolutely (stole / acquired) all the technology they have for solar, EV, and grid based storage. They have literally innovated 0% in this particular industry. I don’t think there’s any debating this aspect.

          At the same time, China has pour billions into domestic production of solar panels, lithium and sodium batteries, vehicle production, and grid based storage solutions the likes that no other country has even remotely attempted. They recent demonstrated cheap sodium based 10MWh storage systems that can be built using seawater sodium. Something that California makes a shit ton of in their desalination plants, that they currently just shove the salt off as waste byproduct.

          Like, if we wanted to, that kind of thing that China just demonstrated, we could be building GWh level storage systems for 10% the cost of a 1 GWh nuclear facility strictly off a byproduct that California distinctly doesn’t want and is literally paying people to take away. They could literally flip a cost into a revenue stream, but we don’t because “reasons”. We could literally have large batteries charged in Utah, and then use rail to move the sodium based batteries into the Eastern sections of the US, using literally the same infrastructure that we use today to move the tons of coal we move around for the TWh of power we generate. We could be doing this today. But we don’t because many nations just buy the arguments politicians feed them, or “it’s complicated”. And then there’s China demonstrating at small scale that it’s doable. So instead we say “oh well it wouldn’t scale” or “oh well you stole all that tech” because apparently our pride is more important than climate change.

          The thing is, yes China has not committed to educating their population into novel development of these technologies. But at the same time they are deploying this stuff at rates every other developed nation has said they’d like to try and do that one day off in the future. Or can’t do right now because their hands are tied.

          For the folks pointing at China as the enemy, fine. I’m not going to debate it. But there’s still things to learn from what they are doing with that stolen technology. Do we need to cozy up to them? Nah. But they’re showing off that grid based storage at scale and cheap is a thing even though people like France and the US say that such a thing is not possible at this time. They are showing LFP is viable if you’re willing to take an initial domestic loss to invest in the infrastructure, something the US citizens know but keep saying “well oil interest are holding us back”. No, there’s only a few dozen oil execs, there over a three hundred million non-oil execs. It’s a lack of will power.

          Like most western nations keep coming up with excuses for delaying EV and green technology pushes and China keeps showing many of the excuses given to be false. And we know they’re false. We know the expectation of no less than $36k USD for an EV is some bullshit that car companies are pulling to offset all the baggage they have from leaving ICE. We know we could have charge stations every 100 miles on the Interstates, but we don’t because oil companies don’t want to lose their investments in the infrastructure they’ve got right now.

          We know the reasons being given by our political and industry leaders are all bullshit. China is over there showing IRL how bullshit they are. Yeah, they stole everything they have, but at the same time all this “oh we couldn’t possibly do that here in the US” is shown for the BS it is, that we already know it to be, in China.

          I mean, great, we’re all very smart people. Awesome. What good is that awesome smartness if we keep letting dumb fucks in politics pander off dumb excuses for why we don’t get to enjoy any of the stuff that awesome smartness provides? What good is being innovative if corporations keep handicapping that innovation to ensure they have a steady stream of revenue?

          I mean yeah, let’s call China out of the bullshit they pull. But I mean, let’s not forget all the damn windows we’ve broken ourselves in our glass house here.

          • bufalo1973@lemmy.ml
            link
            fedilink
            English
            arrow-up
            1
            ·
            6 days ago

            Why move the batteries instead of “moving” the electrons? You generate the electricity anywhere you want and use Therese nice cables that happen to be everywhere.

          • foofiepie@lemmy.world
            link
            fedilink
            English
            arrow-up
            3
            ·
            7 days ago

            Just my take but:

            Like them or not (and IMV they are a serious threat), China’s system enforces a strategic view, long term, more like a 100yr plan.

            We don’t. It’s by election cycle or quarterly earnings report.

            These things all make more sense if you see them impassionately, and without an ethical filter, from a long term POV.

            China will do what’s best for China in the long term. Irrespective of ‘politics’ that are like ripples upon a rising tide.

        • sunzu@kbin.run
          link
          fedilink
          arrow-up
          2
          arrow-down
          4
          ·
          7 days ago

          That’s called value investing… Maybe our dear leader should learn how to manage national wealth instead of cutting companies and allowing a geopolitical adversary to take over tech/IP

          Ie this is not a flex you think it is, it just proves my point that our dear leaders are incompetent imbiciles or worst… Bad faith actors.

          No accountability leads to this sort of decision making lol

      • TimeSquirrel@kbin.melroy.org
        link
        fedilink
        arrow-up
        1
        arrow-down
        1
        ·
        7 days ago

        I’ve seen what’s inside the speed controllers and battery monitoring circuitry for Chinese EVs. I don’t think I want to be anywhere near them.

  • A1kmm@lemmy.amxl.com
    link
    fedilink
    English
    arrow-up
    14
    ·
    6 days ago

    GitHub are not some bastion of righteousness - they are literally owned by Microsoft. And they work hard to stop people from getting too much Open Source from them, with rate limits and the like, so essentially gate keep.

    I think CSDN probably want to gatekeep their clone even harder, but in general having archives of GitHub on the Internet is a good thing.

  • csm10495@sh.itjust.works
    link
    fedilink
    English
    arrow-up
    22
    ·
    7 days ago

    It’s a bit odd, but isn’t it equivalent to forking and putting up a fork elsewhere?

    I guess I don’t see the problem.

    • UnderpantsWeevil@lemmy.world
      link
      fedilink
      English
      arrow-up
      9
      ·
      7 days ago

      It will be funny to see folks who spent the last ten years posting “It’s not stealing, it’s copying” memes suddenly find religion because Evil Foreign People got involved.

    • pumpkinseedoil@sh.itjust.works
      link
      fedilink
      English
      arrow-up
      2
      ·
      6 days ago

      The only issue I see is that they make a new Chinese equivalent for GitHub where they can censor code easier (or was GitHub already blocked?), but they already censor everything anyway so there’s probably effectively no change.

      • SorteKanin@feddit.dk
        link
        fedilink
        English
        arrow-up
        2
        ·
        6 days ago

        Does it though? You can still put up a fork somewhere else as long as you uphold the license right? Unless I guess in the case where the license explicitly disallows forks, but I don’t think that’s very common (can you even do that?).

        • dev_null@lemmy.ml
          link
          fedilink
          English
          arrow-up
          2
          ·
          6 days ago

          Most GitHub repos don’t have a license, meaning you are not licensed to do anything with them. Rehosting them would be the same as rehosting an image you don’t have a license for.

        • barsoap@lemm.ee
          link
          fedilink
          English
          arrow-up
          1
          ·
          6 days ago

          Forks are derivative works (quite obviously) so yes you can forbid them via license terms. Whether or not that’s still open source, take it up with OSI. I vaguely recall that at least once upon a time there was some project that required modification to the code to be published as separate patches and it was generally accepted to be open source don’t ask me which.

    • WanderingVentra@lemm.ee
      link
      fedilink
      English
      arrow-up
      2
      arrow-down
      1
      ·
      7 days ago

      Ya, I kind of like the idea of code being put somewhere else just in case. It sucks it’s China, but I hate to see anything centralized in one company, especially if it’s a big public, good like Github and all it’s code.

  • mariusafa@lemmy.sdf.org
    link
    fedilink
    English
    arrow-up
    2
    ·
    7 days ago

    Well now chinese companies that use free softwware don’t have an excuse to share their modifications of their software product.

  • Muffi@programming.dev
    link
    fedilink
    English
    arrow-up
    10
    ·
    6 days ago

    Great! Now I know who to contact when I accidentally delete all the plaintext API keys and passwords I had stored in a public github repo.

    • OsrsNeedsF2P@lemmy.ml
      link
      fedilink
      English
      arrow-up
      1
      ·
      5 days ago

      Apart from the dozens of scrape bots that already stole them?

      You’re supposed to revoke API keys that are leaked. Not try to “unleak” them

  • Freuks@lemmy.ml
    link
    fedilink
    English
    arrow-up
    2
    ·
    2 days ago

    China cares of nothing, from patents to licences. Culture of steal and copy, rebrand and sell/use

    • rottingleaf@lemmy.zip
      link
      fedilink
      English
      arrow-up
      1
      ·
      2 days ago

      I would argue that this culture would possibly be good to learn from them, first. It didn’t come to existence as some kind of social evolution, but was impressed by power.

      Second, at least they are behind Europeans in the culture of genocide.

    • kbin_space_program@kbin.run
      link
      fedilink
      arrow-up
      1
      ·
      7 days ago

      Better to analyze for vulnerabilities. Particularly with a number of governments using open source software hosted on github.

  • bionicjoey@lemmy.ca
    link
    fedilink
    English
    arrow-up
    98
    arrow-down
    2
    ·
    7 days ago

    Solution: create a GitHub repo with Markdown articles outlining human rights abuses by the CCP and have a large number of GitHub users star and fork the repo.

    • Asherah@lemmy.world
      link
      fedilink
      English
      arrow-up
      5
      arrow-down
      12
      ·
      7 days ago

      Maybe we should consider the same for the US government instead of being afraid of the big Chinese boogeyman across the sea? Because I guarantee you the US has just as many, if not more. But China bad. 🙄

      • bionicjoey@lemmy.ca
        link
        fedilink
        English
        arrow-up
        3
        arrow-down
        1
        ·
        7 days ago

        I was making a joke about abusing Chinese censorship in order to stop them cloning GitHub repos (assuming that was something you wanted to do. The joke being that the CCP suppresses information about their human rights abuses. That is not true of the US. You could absolutely make a GitHub repo detailing the crimes of the US government. Nobody will stop you.

      • x4740N@lemm.ee
        link
        fedilink
        English
        arrow-up
        0
        arrow-down
        1
        ·
        6 days ago

        50 Cent Army Repellant:

        六四

        1989 Tiananmen Square Massacre

    • Colonel Panic@lemm.ee
      link
      fedilink
      English
      arrow-up
      57
      ·
      7 days ago

      You’ve heard of CamelCase and lowercase and intVariableName variable naming styles. Get ready for:

      for (int Taiwan == 0; Taiwan < HongKong; Taiwan++) { int TianamenSquare == 0; … }

    • Tramort@programming.dev
      link
      fedilink
      English
      arrow-up
      28
      ·
      7 days ago

      That’s the whole point of this: they will automatically filter that out, and this is an impotent, though well intended, gesture.

      • bionicjoey@lemmy.ca
        link
        fedilink
        English
        arrow-up
        7
        ·
        7 days ago

        Yeah I figured as much. It was mostly a joke. At the end of the day, if stuff is on GH, people can take it. It’s barely even stealing. Unless the license disagrees of course but then you were putting a lot of trust in society by making it public in the first place.

        • jaybone@lemmy.world
          link
          fedilink
          English
          arrow-up
          2
          arrow-down
          1
          ·
          7 days ago

          That’s what I don’t get about this. Why does anyone care? Even this Chinese company, why do they care to clone it all? It’s already all hosted and publicly available.

          • irreticent@lemmy.world
            link
            fedilink
            English
            arrow-up
            2
            ·
            5 days ago

            Even this Chinese company, why do they care to clone it all? It’s already all hosted and publicly available.

            Until it isn’t. Perhaps they are preparing for a future war with the US and assume their access to all that code will be blocked. They want to copy it now while they have access.

          • bionicjoey@lemmy.ca
            link
            fedilink
            English
            arrow-up
            6
            ·
            7 days ago

            Apparently they aren’t respecting licenses. It’s possible to have source code publicly available on GH but have it not be truly FOSS. But that’s generally not a great idea since you’re effectively relying on the honour system for people not to take your code.

      • Morphit @feddit.uk
        link
        fedilink
        English
        arrow-up
        20
        arrow-down
        1
        ·
        7 days ago

        How will they filter it out? If they just don’t mirror anything with ‘forbidden’ terms, we can poison repos to prevent them being mirrored. If they try to tamper with the repo histories then they’ll end up breaking a load of stuff that relies on consistent git hashes.

        • jorp@lemmy.world
          link
          fedilink
          English
          arrow-up
          6
          ·
          7 days ago

          I feel like the effort to make such a repo and make it popular enough to be cloned and rehosted is a lot more effort than someone manually checking the results of an automated filter process.

          The “effort economy” is hugely in favor of the mirroring side

      • Azzu@lemm.ee
        link
        fedilink
        English
        arrow-up
        7
        ·
        7 days ago

        The real solution is to include a few tiananmenSquare variables in all the repositories. Either they exclude the entire repository or just the specific file, in either case the entire project may be unusable.

        • Tramort@programming.dev
          link
          fedilink
          English
          arrow-up
          1
          arrow-down
          1
          ·
          7 days ago

          China filters every byte of Internet traffic in and out of the country.

          It seems naive to think they can’t accomplish the same thing for a GitHub mirror.

          • Azzu@lemm.ee
            link
            fedilink
            English
            arrow-up
            2
            ·
            7 days ago

            They’re not supposed to, it’s just about blocking them from using the software :)

        • BeigeAgenda@lemmy.ca
          link
          fedilink
          English
          arrow-up
          5
          ·
          7 days ago

          It’s a new coding paradigm, I will take some time getting used to looking for libraries in the uyghur/tianamen folder.

    • UnderpantsWeevil@lemmy.world
      link
      fedilink
      English
      arrow-up
      4
      arrow-down
      3
      ·
      7 days ago

      create a GitHub repo with Markdown articles outlining human rights abuses by the CCP

      Once you have logged “China killed 100 Zillion people! End CCP now!” in Chinese GitHub, everyone in China will realize that their lives are actually very bad and they need to do a Revolution immediately.

      • bionicjoey@lemmy.ca
        link
        fedilink
        English
        arrow-up
        21
        arrow-down
        3
        ·
        edit-2
        7 days ago

        Tankie whataboutism strikes again.

        Two things can be bad at the same time. Wild, I know.

        Edit: also, the point of my joke wasn’t the human rights abuses. It is that these things are censored in China. So your comment is even more irrelevant. One could very easily create a repo outlining American crimes and put it on GitHub. But doing so in China with CCP crimes will have you sent to a Gulag

        • sub_ubi@lemmy.ml
          link
          fedilink
          English
          arrow-up
          4
          arrow-down
          21
          ·
          7 days ago

          “Whataboutism” is what Americans say to profess blind faith in their exceptionalism.

            • sub_ubi@lemmy.ml
              link
              fedilink
              English
              arrow-up
              0
              arrow-down
              10
              ·
              7 days ago

              So you have even less reason to use the racist-in-origin and logically fallacious term.

              • bionicjoey@lemmy.ca
                link
                fedilink
                English
                arrow-up
                5
                ·
                7 days ago

                Lmao it’s literally the name of a logical fallacy. How is the term itself fallacious?

                Also I harbour no racism or ill will toward the Chinese people. My girlfriend is Chinese and I care about her a lot and love learning about her culture. I just don’t abide the human rights atrocities committed by any government.