Top
Best
New

Posted by modinfo 5 hours ago

REA Reverse – Engineer Anything(rea.tools)
250 points | 82 comments
InvisibleUp 53 minutes ago|
Glancing at the Touhou 4 decomp[1], it's a lot better quality than a lot of AI decomps I've seen. It's matching, the variables are named sensibly, comments are sparse and comprehensible, and there's not much in the way of unaddressed Ghidra jank. And only in a month! My main complaint is that the file structuring seems more optimized for AI use than for mirroring the original intent of the devs. (Compare to this human-made Touhou 6 decomp.[2])

I know the retro game modding/decomp scene is already getting hit with a flood of low-effort ports.[3] Now they're going to be hit with a flood of half-decent decomps effortlessly generated by anyone with a $200/mo AI subscription, which become half-decent PC ports, which get modding support grafted on. It kinda just... totally eliminates that as a hobby entirely. I know people are going to be upset about this. And it's not even like with the AI art, where the pro and anti-AI camps live in their own camps. Solving the same puzzle twice just feels discouraging. Very similar to the problem academics are having with AI math/physics proofs.

I know a large reason people hate AI is because it's largely seen as tearing apart hobbyist/professional communities, eliminating reasons to collaborate with others and form relationships, and replacing it with an individualized dependence on a commercial product. It follows the same trend the tech industry took social media, from a method of connection to a tool of propaganda and paranoia and "@gork is this true?" This project, I think, is one of the most clear examples of that.

[1] https://github.com/N0zoM1z0/th04 [2] https://github.com/GensokyoClub/th06 [3] https://www.pcgamer.com/gaming-industry/pc-ports-of-old-cons...

mgaldys4 17 minutes ago||
Slightly off topic. I looked at REA's Android reversing support and found it still uses jadx mcp. That kills large scale APK reversing. jadx takes tens of minutes to preprocess an APK, build a code relationship database, etc. Even headless mcp is no exception. I basically can't use it to analyze APKs at scale, like 100 large commercial APKs in a pipeline, or 100 preinstalled APKs from a phone ROM to hunt bugs.

So I built droidasc. No memory bloat, no parsing slowdown. Analysis is in milliseconds. Global xref on a 300MB APK takes 1.5 seconds. I used it with codex to analyze phones from 3 different brands and found 2 RCEs and 5 root bugs in a few days. I plan to detail these at Black Hat Asia 2027. A friend used droidasc to scan various bug bounty targets at scale and found 10+ RCEs. Way, way faster than jadx.

https://github.com/MG1937/ASC

skyfine 7 minutes ago|
Interesting, might be fun to try my hand at it sometime. Is the best way to actually get the APKs to just download them from an apk mirror site.
mgaldys4 4 seconds ago||
Yes! APKMirror is a good choice. AndroZoo works too, but it's academic and you need to apply for API access.
nirav72 3 hours ago||
I wonder if this is why I've seen a lot videos popping up on my youtube feed related to vibe coded clones of various commercial apps in the past few days. Everything from clones of flagship products from Adob to Microsoft Office.

Adobe product clones like Photoshop and Illustrator: https://www.youtube.com/watch?v=eFB79TYI-Vw

Adobe after effects clone: https://www.youtube.com/watch?v=5mi_tYSdkWQ

MS Office suite clone: https://www.youtube.com/watch?v=U_jTYMOlXio

echelon 2 hours ago||
No, we explicitly do not use reverse engineering. We do not decompile binaries, we do everything 100% clean room:

https://github.com/storytold/photocraft (inspired by Photoshop)

https://github.com/storytold/wordcraft (inspired by Word)

https://github.com/storytold/pdfcraft (one of the more mature apps)

https://github.com/storytold/vectorcraft (another app close to 1:1 parity)

(etc.)

Using REA for something as high profile as what we're doing is likely to result in lawsuits. We're doing everything we can by the books.

We cannot look at Adobe sources. Use of Ghidra is disallowed.

REA is probably great for personal apps and for abandonware, but I think if you publish the results and it's found to have decompiled the original proprietary sources in discovery, you might be in for a bad time.

lifeisloving 2 hours ago|||
Using an LLM is not a clean room, imo. Its just IP laundering. Which is fine I guess if everyone is doing it, including the companies you're stealing from. I just dont know what the implications will be for progress.

Licensing/copyrighting encouraged people to think up of new things, and new ways of doing something. Now we're just all copying eachother.

shinyoo 1 hour ago|||
I agree. The vast amount of data ingested and internalized by LLMs has effectively been "laundered". But they are so powerful and evolving so fast that no one can be spared of their impact. We have to to learn to live with it. Traditional proprietary software being "laundered" is just one part of the broader story...
jasomill 1 hour ago||||
Doesn't clean room typically apply to cases where the "dirty" team has legitimate access to copyrighted code, like the IBM PC BIOS which was published in the technical reference manual, and uses this access to write a functional specification for the "clean" team?

I'm not sure how clean room would apply to commercial applications distributed in binary form, as there's no way to look at even disassembled code without violating the license agreement and therefore being in breach of contract and subject to potential copyright infringement claims for copying or even continuing to use the software, let alone cloning it, and surely you're not going to be subject to a copyright claim based on familiarity with the application from merely using it.

20k 57 minutes ago|||
The issue is that LLMs very likely have access to the source code of these products as part of their training data, and are then being used to generate clones. There's no barrier in the middle to ensure copyright violations don't leak

The reason why 'reverse engineering' has gotten so good is because what we're actually seeing is fully automated luxury plagiarism

Geof25 27 minutes ago||
It is very unlikely that LLM had access to original source code of Photoshop.
greggsy 50 minutes ago|||
The practical uses are more applicable to laundering open source code covered by copyleft licensing like GNU.
echelon 2 hours ago|||
This is novel Rust/egui code that I imagine looks nothing like Adobe code. I've never seen their code, but it must be a mess of old C++, right?

It's considerably faster than their apps (at startup) too.

o1o1o1 1 hour ago||||
Please don't take offence by this, but:

If AI models can generate designs faster and produce work that is "good enough", what is the actual future of design tools and the design profession in your opinion?

I've tested this myself with some frontier models and the results are kinda impressive enough that it raised the question if design skills are already obsolete. If that is the case, what use are these tools now?

komali2 1 hour ago|||
Good enough is not really good enough.

I've seen a lot of "really cool unique designs" that are obviously just a digested regurgitation of amalgamated corporate slop. Anyone wanting an actual unique design language needs to hire actual designers.

Those designers may use AI tools but the tooling isn't to where they can be completely replaced. I challenge anyone to show me an e2e LLM design toolkit that can actually replace a designer, not just one shotted "wow that looks so cool" character designs.

sheeshkebab 1 hour ago||||
How do you know llms you are using were not trained on decompiled apps?
Fizz43 1 hour ago||
why would that matter? That sounds like a problem between Adobe and the AI companies.
greggsy 53 minutes ago||||
Would love to see an InDesign clone that supports writing to and from their proprietary format.
idiotsecant 1 hour ago||||
The death of IP in the west. Good thing? Bad thing? Who knows. Definitely the start of something big though.
angusturner 1 hour ago|||
Mightn't be so bad if all the value wasn't being captured by trillion dollar corps, despite them benefiting off the labor and creativity of countless non consenting individuals
geraneum 1 hour ago|||
IP is not dying. It’s being consolidated to just a few mega corps. Good thing? Bad thing? Definitely the start of something big though.
nico 2 hours ago|||
The AI labs already figured it out:

1) reverse engineer the code 2) train a model on the code 3) use the model to write the clean code

Step 2 is the key “cleaning” process

So maybe a good strategy would be to use something like REA, put it on GitHub, wait for the LLMs to train on it, then just use the frontier models

/s

lifeisloving 2 hours ago|||
I prefer the word *laundering
Iolaum 1 hour ago|||
LLM's have already trained on similar enough code
nico 1 hour ago||
Even better, 1 and 2 done, just move to 3 and profit
komali2 53 minutes ago|||
I've been really inspired by it. Ever since that sort of joking project came out, "Malus," that would clean room reimplement GPL software under more permissive licensing, I've been chewing on doing the reverse for proprietary software. I've been calling it the Manfred Macx theory of disruption, for the character in Accelerando that would patent and then release to public domain profitable ideas before a megacorp could lock them down.

I've got my wedding coming up so haven't had the time I want to dedicate seriously to finding some proprietary things to disrupt, but I've had fun using this as an opportunity to play with subagent orchestration, open weight models, various harnesses, and local models, to see what happens if I let some LLMs churn on it. That resulted in a hodge podge researched list of potential targets: https://github.com/508-dev/genairosity/issues

One thing that stands out is a lot of software kinda does already have a FOSS replacement, it's just not really how people want it to be. GIMP being the representative example. Photoshop people just don't like it, I'm sure for not entirely invalid grievances. However the maintainers of these kinds of tools are often strongly opposed to LLM involved contributions, again , often for not entirely invalid reasons on these incredibly complex projects.

So the long and short of it is that as powerful as LLMs are, we aren't quite where some of the doomsayers are saying we are, insomuch as proprietary software is dead. Even with reverse engineering we aren't there. There's genuine labor, time, and expertise moats around most of these programs.

Personally I'm shifting to trying to find niche abandoned software with no export flow. Even if I can only help out a couple hundred people, it sounds like a decent use of my time.

socializer 3 hours ago|||
Probably not, there's relatively little secret sauce to something like Photoshop. It's just a lot of grungy work that, I guess, you can now delegate to an agent if you have enough money and time.

As an aside, I've heard a lot of hot takes about how this is the end of Adobe, but I'm pretty sure it misses the point. The main reason people pay Adobe is because it's a familiar line of stable, well-supported, interoperable, and actively-developed products. There's already plenty of cheaper or free alternatives (Davinci Resolve for video, Capture One / Darktable for raw, Affinity for photo editing and vector drawing, etc), and if Adobe survived that, I sincerely doubt they're going to lose pro customers to a vibecoded app where half the stuff is probably subtly broken or left as a TODO, and that will be abandoned in a week, because the whole point was to get that 1M YouTube views.

echelon 2 hours ago||
> that will be abandoned in a week, because the whole point was to get that 1M YouTube views.

Longbets 1 week, haha.

We're working our asses off on this.

Most of the team are artists who use these tools actively and we want the replacements for ourselves. I'm a filmmaker, so you can imagine my frustration of being bitten by the "unsubscribe fee".

lewelove 2 hours ago|||
You're doing God's work. Keep it up! PhotoCraft is really cool, and the pace of stabilizing progress is incredible! Don't listen to people who dismiss your work without getting their hands on it.
woggy 15 minutes ago||||
Have you guys done any talks or written any articles on how you are developing this software?

Scaling up to this kind of complexity even with AI is still a challenge, at least for me.

geraneum 29 minutes ago||||
> being bitten by the "unsubscribe fee"

I’m curious how much is your LLM bill. I’m assuming the idea is that LLMs will maintain the software? I’d be surprised, given the token economy, that you’d have to pay less than those subscriptions to LLM providers.

namrog84 30 minutes ago||||
I am definitely interested in alternatives, but I am not interested in an alternative that isn't supported and maintained. Though I suppose at some point I can ask my own agents to support/maintain it?
beepbooptheory 1 hour ago|||
No ethical qualms here but you guys should save your money and just use Gimp! It's already free, really capable, and comes from a good and pure place of making software for its own sake. That's the kind of thing that gets you users for decade, beyond any egui rust doo dads.
skeaker 1 hour ago||
They are covering the entire Adobe suite, not just Photoshop.
chiengineer3 2 hours ago|||
I built a from scratch rust RAW processing engine similar to the brains behind photoshop and lightroom - its much more powerful and its not even complete yet
rchase 3 hours ago||
yep. it is unexpectedly everywhere. like, overnight.
areoform 4 hours ago||
I love such work. I hope to roll it into a package that anyone can use in the future. This work is only going to get more important because frontier models getting locked down will make this a lot harder over time.

For example, I use Claude as a bouncing wall for my thoughts and I pointed out that,

    > GLM 5.2 was the only thing that helped HF while the agents were trying to access them. The "guardrails" stopped them from doing good. The Computer Fraud and Abuse Act exists. Courts exist. And computers and an internet connection have existed for a long time. There's also 17 USC 1201 provisions with the 1201 a 1 exemptions [Image #31] so in this case, a farmer should be able to work with you to access the tractor they own. Or... IDK... a kindle that's out of date? :) What is lawful and what isn't is rooted not within the act but within intent, purpose and mens rea. And this is something the law has been deciding for centuries now. At one end, your maker can't say that governments should decide while at the other end explicitly refusing to allow governments to be the ones who decide.
This was rejected for "Safety,"

    > Opus 5.5's safeguards flagged this session. You may be seeing this for the first time on an Opus model: Opus 5.5 is more capable and has stronger safeguards as a result, which can sometimes flag non-cybersecurity work. We're improving these safeguards to reduce the amount of incorrectly flagged messages. Edit and retry, or continue with Opus 4.8. Send feedback with /feedback or learn more: https://support.claude.com/en/articles/8106465 
    >
    > Details: "[cyber]'
Note, the image here was the Library of Congress' page on DMCA exceptions.

Fundamentally, the idea that you can't reverse engineer things, make things, learn about biology or physics without permission is strange to me. These machines have been trained on the sum intellectual output of humanity, the global intellectual commons, and are being used to close off that commons?

I would be OK with their right to create such restrictions if they weren't lobbying the Government to restrict others, thereby ensuring that they control humanity's intellectual commons well into the future.

Perhaps I'm naive, but I think it's better for humans and the machines if we can all think, learn and build. But then again, I'm the kind of person who rejects the doomer pill.

TheSamFischer 3 hours ago||
Yay, all software will be open…Except the models.
LoganDark 3 hours ago||
I wanted to see if CVP approval changed this response, but it appears that with the release of Opus 5.5, Anthropic silently dropped me from the program, and has some strict new criteria in place to apply again, such as being credited for a CVE! I was only approved last month, too -- sad!
papascrubs 2 hours ago|||
I have CVP with the new program (including mythos access) and still get constant denials for silly situations. Most recently I fed a URL to my agent from a security blog and asked if to add it to my obsidian vault with appropriate tags-- cyber flagged. You're not missing much. OpenAI and/or most Chinese models are much more lax in their restrictions.
areoform 2 hours ago|||
This kills the talent pipeline, and it'll create a spam problem for the other folks because now people will try to github PR spam their way to getting on a CVE.

It's worth talking about the fact that you can't even talk about DMCA to a model trained on the Library of Congress unless you're one of the approved people. And that's before reverse engineering something or writing code.

So in this future, it sucks to be you if you're someone trying to make your small app more secure, someone trying to upskill, a tinkerer trying to bypass corporate lockdowns for a device they own (a recognized DMCA exception, btw), a teenager trying to learn about security...

It locks away much of the richness that produced hacker culture behind glass. You can look at their press announcements and PR pieces, but you can't touch.

And as they're lobbying the government for "sensible regulation," this inevitably leads to a future where computing is controlled.

It's the direction their existing reports are taking. They recently released one in September that talked about how they stopped "bioweapons." What were said bioweapons efforts? Oh, it was scientists using Claude for grant writing, paperwork and grammar. At national labs.

These people are basically proud of impeding real research to make better painkillers and study a neglected tropical disease, https://news.ycombinator.com/item?id=49651727

And this is being used to lobby against "dangerous" open-weight models because gasp a scientist might use them to write a grant! To make better antidepressants.

At what point do they start reporting someone taking apart an iPhone and trying to DIY a repair with a schematic as a thwarted "cyber security incident?"

jasomill 1 hour ago|||
A funny, but slightly chilling safety violation I once got was ChatGPT being unwilling to recite the full text of Article I Section 2 of the US Constitution, aborting as soon as it hit the passage about "three fifths of all other persons".

Another funny one was Claude's refusal to provide the original untranslated text of a passage from Dante's Inferno on copyright grounds, though in this case pointing out that no 14th century literature was subject to copyright anywhere in the world was sufficient to override its objection.

LoganDark 1 hour ago|||
Agree on the talent pipeline. It can take a long time for someone to obtain a CVE that has their name on it. You don't start being a security researcher only once that happens.

Several years back, I was working on generating AVB2 hashes on top of modified Android distributions, to increase the security after an owner has made their desired changes. I was doing this before the age of LLMs. Among other things, this would've enabled the secure features to work again, and potentially reduce the risk of root access being usable by malware. But apparently I'm not a security researcher because I didn't get a CVE about it.

WarmWash 51 minutes ago||
I see this liquid software being the unavoidable future. A computer that just does stuff in whatever way you guide it, in whatever way you like guiding it. The models will keep getting better and faster up to the point where everything is just happening in real time, no OS no drivers no programs, just an entity that can listen to you and can move around bits to accomplish whatever you are trying to do. Wanna post on HN? Any way that you can code such an action, the computer can just manifest it for you, on the fly, however you want it.
geraneum 25 minutes ago|
What a dystopian future where your computer use will be basically “metered” by the token. At least right now we don’t have to pay per mouse clicks.
ethin 4 hours ago||
I might be missing something but... How exactly is this better than telling Claude for example to "install and set up a full RE environment including Ghidra" on my local system and get to work? Like what does this do that my current RE methodology doesn't?
none_to_remain 1 hour ago||
I don't have reverse engineering experience, though I have all the prereqs to learn it. Anecdotally a few days ago I told my slow local Qwen3.8 in Pi harness to use Ghidra CLI to decompile a certain executable and it got entirely lost. Today I was linked to this, have it chugging along now, and it's making some sort of progress towards unpacking this thing. I don't know if it will nail it this run but it's a real improvement; I definitely have some useful info I could bring to another prompt or for myself if I cared to try manually.

But I just found an even easier way I should have thought of first - someone already dropped a reimplementation a couple weeks ago.

fwlr 3 hours ago|||
Probably it is not better. I think this is aimed at people who do not have a “current RE methodology”, do not know enough to specify things like Ghidra, etc., but who do have a desire to feel like they reverse-engineered and can reliably predict that a conversation with a chatbot will make them feel that way.
garyfirestorm 3 hours ago|||
Vibe-reverse-engineering
XorNot 2 hours ago|||
I mean thats not why I have Claude decompile things. I have it so it because my software should work exactly how I want it to.
tptacek 4 hours ago|||
Everybody has their own set of skills and specific scripts and tools to do this stuff. You might use Ghidra as the the kernel of those workflows, but you still want something more than just Claude freestyling, at least for now.

(Who knows if this'll be true 6 months from now.)

xpct 1 hour ago|||
I don't know either. My only guess is that this has some helpful context for less powerful models, maybe.

We're kinda far into this LLM thing, maybe it's time to start selling harnesses by leading with how some examples were solved faster with this and how it saved tokens, or similar?

leidenfrost 1 hour ago|||
Why not just, like, ask Claude about it?

"Hey Claude I want to create a fanmade Game Boy game, give me recommendation of tools, libraries and workflows for it" and boom.

You can use AI to learn stuff, instead of just using it as a Pokemon.

ambicapter 1 hour ago|||
Because it's already been done for you? Sure, you can spend your own tokens on it, but it's probably cheaper and more time-effective to use something that already exists and does the job.
djmips 4 hours ago|||
I don't think it is better than just doing it yourself in Claude Code (in fact worse) but some people like cute UI for everything I guess.
echelon 2 hours ago||
We're using vanilla Claude Code for ArtCraft apps [1], but we are especially careful not to touch Ghidra. We don't want decompilations or reverse engineering to spoil the work we're doing and expose us to copyright infringement.

[1] https://github.com/storytold

methou 1 hour ago||
I'm using IDA Pro's MCP (Well worth of money in the past), it may hit the infamous artificial cyber wall at any time.

Try GLM-5.3, it worked pretty for me. It also worked well with Radare2 or binary ninja if you don't have the muscle memory for idapro.

jasomill 1 hour ago|
This makes no sense to me. There are innumerable reasons to reverse engineer software that have nothing to do with security, some not only not prohibited, but in fact specifically authorized by law.

It'd make more sense to reject reverse engineering commercial software on the grounds that it likely violates the software's license agreement, and would therefore subject the user to potential breach of contract and copyright claims.

But this would also apply to uploading pretty much any non-self authored document to the LLM in the first place, albeit with fair use as a possible defense after the fact, so it still doesn't make much sense.

Meleagris 1 hour ago||
I’m surprised people don’t get refusals running this, or are they running it with cyber-enabled models like daybreak?

In my experience, even just hinting at reverse engineering to models from Anthropic or OpenAI leaves them extremely sensitive to refusals. After all, the same techniques used here can be used to find exploits.

Or are people using it with local models?

Curious.

wpm 1 hour ago||
I handed Astra Ghidra with some printer driver exes/dlls loaded, and a Windows 2000 VM with the software installed and said "reverse engineer this printer driver, it's hooked up and powered on on /dev/ttyUSB0, tell me when you're ready to print a test page". No refusals.
Scaevolus 1 hour ago|||
Reverse engineering and decompilation is not currently blocked by the models, leading to the proliferation of LLM-assisted decomps/recomps. If you ask it to find bugs you're more likely to get a refusal, but simply pointing an LLM at a Ghidra MCP server and telling it to trace program flow, rename functions, or answer questions behaves like normal.
marcelo-earth 1 hour ago|||
Yes, it just rejected me once with Claude Code Opus 5.5, but it works fine with Codex 6.1-Sol (in my experience, it was the other way around)
xpct 1 hour ago||
I think people said that Claude refuses to work on RE sometimes.

I've had no issues with OpenAI.

armcat 2 hours ago|
I’ve been trying to rebuild and modernise an old abandoned ms dos game, simply by pointing Codex (6.1 Sol) at the game directory, and it works insanely well. It’s able to understand the data formats, unpack graphics and sound assets, and reconstruct game logic.
More comments...