GitCode, a git-hosting website operated Chongqing Open-Source Co-Creation Technology Co Ltd and with technical support from CSDN and Huawei Cloud.
It is being reported that many users’ repository are being cloned and re-hosted on GitCode without explicit authorization.
There is also a thread on Ycombinator (archived link)
Solution: create a GitHub repo with Markdown articles outlining human rights abuses by the CCP and have a large number of GitHub users star and fork the repo.
Maybe we should consider the same for the US government instead of being afraid of the big Chinese boogeyman across the sea? Because I guarantee you the US has just as many, if not more. But China bad. 🙄
I was making a joke about abusing Chinese censorship in order to stop them cloning GitHub repos (assuming that was something you wanted to do). The joke being that the CCP suppresses information about their human rights abuses. That is not true of the US. You could absolutely make a GitHub repo detailing the crimes of the US government. Nobody will stop you.
Tell that to Julian Assange
Is that what you think got him in trouble?
deleted by creator
With the obligatory “fuck everyone who disregards open source licenses”, I am still slightly amused at this raising eyebrows while nearly no one is complaining about MS using github to train their copilot LLM, which will help circumvent licenses & copyrights by the bazillion.
If I look at a few implementations of an algorithm and then implement my own using those as inspiration, am I breaking copyright law and circumventing licenses?
That depends on how similar your resulting algorithm is to the sources you were “inspired” by. You’re probably fine if you’re not copying verbatim and your code just ends up looking similar because that’s how solutions are generally structured, but there absolutely are limits there.
If you’re trying to rewrite something into another license, you’ll need to be a lot more careful.
What’s the limit? This needs to be absolutely explicit and easy to understand because this is what LLMs are doing. They take hundreds of thousands of similar algorithms and they create an amalgamation of it.
When is it copying and when it is “inspiration”? What’s the line between learning and copying?
I disagree that it needs to be explicit. The current law is the fair use doctrine, which generally has more to do with the intended use than specific amounts of the text/media. The point is that humans should know where that limit is and when they’ve crossed it, with motive being a huge part of it.
I think machines and algorithms should have to abide by a much narrower understanding of “fair use” because they don’t have motive or the ability to Intuit when they’ve crossed the line. So scraping copyrighted works to produce an LLM should probably generally be illegal, imo.
That said, our current copyright system is busted and desperately needs reform. We should be limiting copyright to 14 years (as in the original copyright act of 1790), with an option to explicitly extend for another 14 years. That way LLMs can scrape comment published >28 years ago with no concerns, and most content produced >14 years (esp. forums and social media where copyright extension is incredibly unlikely). That would be reasonable IMO and sidestep most of the issues people have with LLMs.
Some random Chinese company: does something jenky
Blogger: “The entire country of China is doing this jenky thing!”
This comment should be deleted soon
This is inevitable:
Once the people in China can only see the CCP’s version of everything,
& ALL stuff has been adulterated, either by AI or by some agency-or-other,
THEN dissent should die-down in the Chinese population:
Read Lanier’s “Foreign to Familiar” to understand how Tropical-Culture vs Nordic-Culture shapes people, & how old-cultures vs new-cultures shape people,
then read Hofstede’s “Exploring Culture” to understand the dimensions of culture that his Cultural Dimensions Theory digs into ( power-distance, uncertainty-avoidance, “success”-orientation, & other dimensions )…
& when you understand how we’re kind of “template” people, before being born into culture,
but once born into it, our entire meaning gets framed within whatever culture we were born into…
therefore, the CCP can simply remove most diversity-of-meaning from their completely-possessed-population, through a generation or 2 of that.
Tibetan, Uyghur, Hongkonger, Taiwanese, Indian, South-Korean, Japanese, the intent is consistent: "the destruction of " … others … “is the midwife of Chinese supremacy”.
I expect a similar kind of program to exist in all right-possessed countries, as the right is doing in the US, right now, with burning or banning books, eradicating proper education, suppressing libraries, etc, they’re just doing the same thing as what the CCP’s doing, only less-skillfully, is all.
No real difference in their deeper heart/motivation/intent, though: supremacism, crushing/destroying all “other” kinds.
Russia’s big on it, too, isn’t it?
Islamism…
The “Crusades” were good examples of this kind of idiocy?
The “Inquisition”?
The “Buddhist” genociding of Tamils?
So long as the “home” story is … “coherent”, & “justifies” all, then … kids grow up … believing, right?
There’s a book, & a Big Think yt video, on “Collective Illusions”, which is important!
Please invest in seeing that video, & see how it’s actually a delusion-mechanism in our minds…
…used by political-forces, yes, but they couldn’t use it if it didn’t exist, could they?
_ /\ _





