FRESH

Hacker News

Home

Ask HN: Is "no source code was copied" still a sufficient copyright defense?

64 points by oscgam1

by arjie

5 subcomments

The Corgi event doesn't seem particularly notable. There are similar features implemented in the most bog standard way that those features can be implemented using the pattern that AFAIK Github pioneered with a 'Danger Zone'. Both parties are using the same upstream components so it ends up looking the same.
I don't know when the extreme intellectual property viewpoint entered software engineering as a mainstream opinion because I have never before seen it expressed so strongly in this community (seeing as I wasn't around when Bill Gates famously asked for money first or whatever). To think that a past OpenOffice would have been considered unconscionably close to a copy of an old MS Office of the era twenty years ago.
In some way, The Corporations Won, because it turns out software engineers turned into IP maximalists. Thinking back to when I first installed Tux Kart decades ago I never could have imagined that we'd get to this stage. Really wild, man.

by otekengineering

1 subcomments

OpenAI/et al. selling an IP laundering service under the name 'max subscription' may force the world to accept the perspective that Intellectual Property isn't a thing. The business model of extracting value from creators via rent seeking IP may not be viable in a world where LLMs can generate anything on demand. We might be transitioning to the Lockean view that for something to be ownable as property, it must be a scarce resource, and information is not a scarce resource.
From that property rights perspective, the property that's created when new information is created is not the information itself, rather, it's the act of creation (claim to authorship) that's the scarce resource.
I don't know what a world looks like where the only form of IP is non-transferable and owned by the original creator. Maybe that new form of IP creates less value over all, and maybe that's ok if the creator is getting 100% of the smaller pie instead of crumbs from media labels. Companies like Red Hat could be an example of a viable business model if IP laws follow the current winds.
Companies like Corgi will need to rely on internal talent to ensure that their product is better than what someone looking at their product can vibe code a copy of, which from my perspective as a consumer, sounds like a better route than Corgi relying on an internal legal team to send a cease and desist letter.

by egypturnash

1 subcomments

"Now software developers are feeling what authors and artist felt".
As an artist who got repeatedly told to stop making buggy whips and get into the absolutely tedious-sounding new field of "writing prompts" every time I expressed dismay and displeasure about image generation around here, every story about this sort of thing here is the sweetest schadenfreude I have tasted in my life.
Especially when the general feeling in the markets I work in is that AI images are kinda tacky and empty and nasty, and people would rather pay another human to realize their ideas than try to refine image generation prompts for a couple hours and get something vaguely okay that makes people go "ew, AI".

by dlcarrier

1 subcomments

Copyright doesn't cover instructions like recipes, protocols, or APIs; those require patents.
Not looking at the source code has been used to make nuisance copyright lawsuits less likely (e.g. Phoenix and AMI implementations of IBM's BIOS) but it's still easy to prevail when a new work is created by rewriting some else's source code. (https://en.wikipedia.org/wiki/UNIX_System_Laboratories,_Inc.....)
Neither copyright nor patent cover a user interface (https://en.wikipedia.org/wiki/Apple_Computer,_Inc._v._Micros....), so that can legally be copied outright.

by aldousd666

2 subcomments

Copyright doesn't cover the results of code, nor the methods used in the code, techniques and algorithms aren't covered by copyright. Period. Copyright applies to 'the work'. If you don't copy the source code, it's not covered.

by eqvinox

3 subcomments

"still"? It never was. If you copy a (copyrighted) UI in bulk, that's a copyright violation just like copying code in bulk. The legal metric is generally "sufficient height of creation", the actual interpretation depends on where you are.

by wahern

0 subcomment

If you're worried about infringement, register your work with the US copyright office. You can only get monetary and statutory damages if the work was registered before infringement, otherwise you can only get an injunction. But you can't even file a claim in court to request an injunction without first registering the work. Basically, while copyright nominally attaches at creation, without a certificate you can't press any rights in court.
You don't need to register each release, so long as a material portion of the registered work exists in subsequent derivative works.
Without a registration threats of a copyright dispute are mostly noise to someone savvy enough to know how the game is played. If they think you'll persist they can just replace the infringing work or cease distribution, which is a hassle but not a significant deterrence for bad faith actors.

by glimshe

2 subcomments

Software copyrights are among humanity's worst inventions. We as a species are no better off because of it, and neither are the small creators that copyrights are supposed to protect. Software copyrights only exist to protect a renter model from big corporations.
There's an argument to be made for patent protections, but many of those are questionable considering the number of trivial software-related patents (there must be a patent somewhere for replying to an online conversation through an edit box and an "add comment" button).
I don't know if LLMs can somehow help the situation. I hope they can expose the ridiculousness of software copyrights but I won't be holding my breath.

by davebren

0 subcomment

In effect the source code is being copied by the LLM. This is what it's designed to do. LLMs are a lossy statistical compression of their training data.
If you give it a prompt telling it to replicate a product that's in its training set then its optimal next token prediction output is going to be to a lossy copy of that product's source code.

by mrdependable

0 subcomment

A sad state of affairs when the law is what people look to in order to decide between right and wrong.

by stronglikedan

1 subcomments

There are no novel UIs, so copying UIs is okay, and necessary. As for source code, I'm a stickler for the license. The modern set of licenses cover any scenario I can think of, relatively fairly. AI is merely a tool, so the craftsman still owns the output. If the output violates a license, then the craftsman should be held to account.

by axus

0 subcomment

I'd argue that software is an "applied art", and needs a "high threshold of originality" to be protected.
https://en.wikipedia.org/wiki/Threshold_of_originality
Oh and if it's not human generated, you can just copy it.

by sandeepkd

0 subcomment

I think there are couple things going on here
The replication/copying has always been there in one form or another. The bar has traditionally been higher for reputation and monetary risks.
Lately the legal bar is the one that going down, ease of replication makes it even more tempting and when big players are doing it at scale (bots) then it validates the strategy in one way or another.
If anything, there have to be downstream consequences of this with time, libraries to pollute the front end code for LLMs are most likely going to get popular and probably one way to make it harder for your IP to be replicated.

by Havoc

1 subcomments

Think it'll be hard to define in any sufficiently specific manner legally because UI/UX/source/functionality aren't entirely separate, but ethically I reckon it comes down to what one means by copy UI.
e.g. If you're creating an uptime dashboard...they all kinda look the same anyway and there aren't that many ways to do it so that seems OK. If it's copying an comprehensive UI with layout and flow between the various pages etc then you're getting a bit closer to theft.

by robotmaxtron

1 subcomments

by conartist6

0 subcomment

So to be clear the answer is emphatically "no". If you copy everything else, the defense that the source code is technically different will not save you.

by 5701652400

0 subcomment

copying other business pixel-to-pixel and direct claims agains competitor is just distasteful. add to this recent YC "AI that MITM API and re-implments anything automatically" is very bad image to YC.
with such bad behavior from SWE community, you just got to lock down your app behind certificate pinning, hardware attestation, gRPC/protobufs, and internal data only. no more "free open web in browsers" when you get gents like this stealing other peoples efforts.

by throwaway81523

0 subcomment

The clean room PC compatible BIOS's were written that way for a reason.

by 8note

0 subcomment

we need open training data models, and not just open weights.
if you cant prove that the source code wasnt trained on, how can you show that its not a copy of the copywrited original?

by 5701652400

0 subcomment

technically may be it is. practically you are a dick.
(e.g. instagram copying snapchat)

by erelong

0 subcomment

...hopefully as other comments said, that LLMs make us abandon the confusing idea of "intellectual property" so we don't have to ask questions like this or get in to torturous questions if this or that thing is "infringing" or not

by crest

0 subcomment

By that logic OpenOffice would infringe on Microslop Office because "it looks the same" (as an older version of M$ Office).

by tamimio

0 subcomment

There was an article here about stealing the other day and only change 3%, so I guess it’s working already!

by jasonlotito

0 subcomment

It depends. Is the text auto-generated from a framework? Is it creative? Is it worthy of copyright? It seems more instructional.
I say let them sue for copyright infringement for the text. Let them sue for breaking a license. Let's see how it works out. Doesn't impact me and if it costs the rich money, good. Let them suffer.

by josefritzishere

1 subcomments

No definitely not. I've never seen a patent include code. They're more likely to describe IP in a work flow diagram.

by kmeisthax

1 subcomments

There's a couple of related issues being conflated here, and I'm not sure which one to bring up, mainly because I'm not sure in what direction the copying would be ruled to have gone. So I'll just mention all the cases.
The first thing to note is that nonliteral copying can still be infringing. Actually, among copyright cases that actually go to trial, most of them are not bit-exact matches ("striking similarity" in legalese). The lower standard those cases would have to meet is substantial similarity, which requires proving both access and similarity. In other words, in order for you to produce a copy[0], you have to have both seen the original and produced something that is close enough if you squint at it.
So let's say Papermark is the original and Corgi Dataroom copied it. If this actually went to trial, a significant amount of discovery would be spent harvesting all the e-mails and messages Corgi's development team sent to one another. Any evidence of knowledge or access to Papermark would probably be enough to prove a copyright violation.
You mentioned LLMs, and some of the tweets here also mention them. I have no clue if either product used an LLM, but it's important to note that in the US, anything written by an LLM does not accrue copyright protection. So, if Papermark was LLM-authored, as a threshold matter, they would have to register[2] a very specific copyright that neatly carves out the LLM-authored bits. The judge would then only consider the parts of the code with verifiable human authorship, which would severely weaken Papermark's case.
In the reverse case - i.e. Corgi Dataroom is LLM-authored - then Papermark's case becomes stronger. Note how I didn't say "AI slop is public domain" last paragraph, because it isn't. There are still unsettled legal questions as to whether or not training on copyrighted works is legal and if using an LLM trained on that work constitutes access to it. Furthermore, LLMs can have search tools that would give them access to code not within the training set, which would also be a more straightforward copyright violation. So if it turns out Dataroom's developers are all Claude fans, and Claude is copying Papermark code, then it's just a normal license violation. LLMs do not launder copyright.
We can also consider the case where BOTH tools are LLM-authored. In this case, there might just not be a copyright case at all. Technically speaking, there would be some third class of plaintiffs who have been infringed, but they would have to choose to sue. It is a long-standing principle in law that you are not allowed to sue for other people's harms[1]. So in this case, nobody would have a case.
> Now software developers are feeling what authors and artist felt
It was specifically the FOSS community that sounded the alarm about training data theft first, because the FOSS community correctly understood coding agents to be an attack on copylefts[3] (albeit through the incorrect belief that LLMs could launder away copyright interest, as opposed to it just being told to make a noninfringing substitute).
[0] I am skipping over notions of fair use and derivative works as they would complicate the analysis and do not apply here.
[1] In general, the operating principle of American courts is "fuck around and find out", and this implies that the court is only allowed to find out once someone has fucked around. Otherwise, the courts could just sue themselves to rule on whatever the hell they want. Isn't adversarial common law GREAT!?
[2] Yes, copyright registration is mandatory in the US, otherwise you can't sue, which is the whole point of copyright. The Berne Convention only half-applies here.
[3] Clauses in licenses that require modifications to the work to be provided under the same license terms. Creative Commons calls these Share-Alike licenses.

by yieldcrv

0 subcomment

not a copyright issue
there may some other intellectual property remedy, or not, but it isn't copyright
hope that helps

by jessebradner1

0 subcomment

[dead]

by mrkimsh

0 subcomment

[flagged]

by negergreger

0 subcomment

[dead]

by dataviz1000

1 subcomments

Did you agree to terms of use? Did you have to click a check box that you agree to terms of use before seeing or having access to the items you copied? Click wrap. If in the contract that you agreed to there is language that you agreed to not copy the work, then you likely are in breach of contract. If it is publicly available knowledge probably not breach of contract. I’m not a lawyer of course.