메뉴
HN
Hacker News • 11일 전

F-Droid 앱 중 얼마나 AI가 만들었을까?

IMP
5/10
핵심 요약

FOSS 앱 개발자이자 학생인 필자가 오픈소스 안드로이드 앱 스토어 F-Droid에 AI(LLM)가 생성한 '바이브 코딩' 앱이 얼마나 많은지 분석한 글입니다. 완전히 판별하는 것은 불가능하지만 AI 생성 아이콘, 성의 없는 README, 코드 리뷰 부재 등의 징후로 판단할 수 있다고 설명합니다.

번역된 본문

목차: 서론 / 문제점 / 바이브 코딩 앱인지 어떻게 알 수 있는가? / 나의 편견 / 기준 및 실험 / 앱 목록(Amber, Aria for Misskey, Atmo Engine, Aves Libre, Balance, Baly Groceries Tracker, Bati: Fitness RPG, BayesianBahn, BeatBridge, Binary Eye, BlockDrop, Braincup, BVD, CaptureCap, Casio G-Shock Smart Sync, Chompass, DeltaSync, DuressKeyboard, eQuran, evcc, Fechtkarte, Feeder, Felicity Music Player, FixupXer, Flexify, Forkyz, Gem Wallet, GeoWeather, GitHub Trending & Hacker News, GPTMobile, GymMane, Harp, Hue Spill, invoice/quote/delivery note, KeyStoreViewer, kitshn, Klick'r, Lens HRV, Léon, LetterBox, Libre Contacts Backup, Lissen, MainTask, Mako, MarketMonk, Markleaf, MateDroid, Materialious, MetaGer, Mihrab, Minesweeper, MinkLauncher, motd, N-Zik, neutriNote CE, Nextcloud, Nextcloud Pantry, NFC Alarm Clock, NouTube, NovaDial, Offline Translator, OPNsense Manager, PCAPdroid, Personal Stuff, Phylax, PingOff, PipePipe, PlainApp, Privacy Flip, Queens, ReadBear, Relatrix, ReteClock, ReteKey, Share To InputStick, Shattered Pixel Dungeon, ShelfDroid, ShizuWall, Simple Notes Sync, Sky Map, Snowdrop, Suntimes Calendars, TacticMaster, Tallybook, Tasks.org, Terminator, TigerDuck, TimeLimit.io, Timety, trale, Tuisku, Unciv, Universal Installer, UnlicenseLauncher, Victoria Launcher, Voxscribe, Water Sort, WaveUp, Wikipedia, Wristotle Companion, Xime IME, Yubico Authenticator 등) / 흥미로운 관찰 / brandonp2412 / Codeberg / "건드리지 마라" 앱들 / 통계 및 결론

나는 F-Droid를 사랑한다. F-Droid가 상징하는 가치를 사랑하고, 사용자에게 주는 자유를 좋아한다. FOSS 앱 관리자로서 이 프로젝트 뒤의 사람들에 대해서도 칭찬밖에 할 말이 없다. 내가 빌드 전에 커밋하는 걸 깜빡해서 reproducible build가 몇 번이나 실패했는데도 친절하게 대해줬으니, 어쩌면 너무 착할지도 모른다 😅. 하지만 정말이지, 요즘은 인간이 직접 작성한 소프트웨어를 찾기가 어렵다.

서론

문제점: 가끔 그냥 구경하려고 F-Droid를 연다. 몰랐던 문제를 해결해주는 앱을 발견하거나, 이미 쓰는 것의 더 나은 대안을 찾을 수도 있으니까. 그러던 어느 날 둘러보다가 뻔하고 흉한 AI 생성 아이콘을 가진 앱을 발견했다. 그 앱은 목록에 없고 지적하지도 않겠지만, 궁금해졌다. F-Droid의 얼마나 많은 부분이 AI로 만들어졌을까?

재미로 프로그래밍을 하는 학생에 불과하고 정직한 노동으로 물들지 않은 나로서는, 코딩 세계가 어떤지에 대한 지식을 대부분 자극적인 유튜브 영상과 CS 전문가들의 레딧 게시물에서 얻는다. 그들은 LLM을 신의 화신처럼 묘사하거나 아니면 고급 자동완성 정도로 묘사한다. 둘 다 분명히 틀렸지만, 내가 알고 싶은 것을 파악하는 데는 그다지 도움이 되지 않는다. 요즘 프로그래밍의 실제 상태는 어떤가? 내가 FOSS 소프트웨어를 사용할 때, 그것이 어느 날 오후에 누군가 대충 바이브 코딩한 것일 확률은 얼마나 될까?

바이브 코딩 앱인지 어떻게 알 수 있는가?

요점은—알 수 없다는 것이다. 텍스트에는 어떤 평가라도 정확하게 만들 만큼의 메타 정보가 충분히 담겨 있지 않다. 하지만 내가 첫 문장에서 쓴 em-dash(—)가 아마 여러분 뇌에서 경보를 울렸듯이, 징후는 존재한다. 이렇게 들으면 이 작업이 변덕스럽고 우연에 좌우되는 것처럼 들리겠지만, 특히 LLM 코드 저장소의 경우 징후는 결코 찾기 어렵지 않다.

LLM의 주된 매력은 개발자를 더 게으르게 만든다는 것이다. 그게 사실 핵심이다! 프롬프트만 입력하고 앉아서 편하게 있으면 된다. 그래서 이런 태도가 바이브 코더가 손대는 모든 것에 반영된다는 것을 들어도 놀라지 않을 것이다. README를 직접 쓸 필요가 있나? 웃기고, LLM한테 시키면 되지. 코드 리뷰를 할까? 됐고, LLM이 자기 변경사항을 스스로 리뷰하게 두자.

원문 보기
원문 보기 (영어)
Table of Contents Intro The problem How to know if an app is vibe-coded? My biases Criteria / The experiment Apps Amber Aria for Misskey Atmo Engine Aves Libre Balance Baly Groceries Tracker Bati: Fitness RPG BayesianBahn BeatBridge: Bluetooth Music Binary Eye BlockDrop: Block Puzzle Braincup - Brain Training BVD CaptureCap Casio G-Shock Smart Sync Chompass - Calorie Tracker DeltaSync DuressKeyboard eQuran evcc - solar charging Fechtkarte Feeder Felicity Music Player FixupXer - URL Enhancer Flexify: Gym Workout Log Forkyz Gem Wallet: Bitcoin, USDT, BNB GeoWeather GitHub Trending & Hacker News GPTMobile GymMane Harp Hue Spill invoice, quote, delivery note KeyStoreViewer kitshn (for Tandoor) Klick'r - Smart AutoClicker Lens HRV Léon – The URL Cleaner LetterBox Libre Contacts Backup Lissen: Audiobookshelf client MainTask Mako MarketMonk: Stock Tracker Markleaf MateDroid Materialious MetaGer: Search & Browser Mihrab: Prayer Times & Quran Minesweeper MinkLauncher OpenSource motd N-Zik neutriNote CE Nextcloud Nextcloud Pantry NFC Alarm Clock NouTube NovaDial Offline Translator OPNsense Manager PCAPdroid Personal Stuff Phylax PingOff PipePipe PlainApp: Phone Web Portal Privacy Flip Queens ReadBear Relatrix ReteClock ReteKey Share To InputStick Shattered Pixel Dungeon ShelfDroid ShizuWall Simple Notes Sync Sky Map Snowdrop Suntimes Calendars TacticMaster Tallybook Tasks.org: Open-source To-Do Lists & Reminders Terminator TigerDuck TimeLimit.io Timety trale Tuisku Unciv Universal Installer UnlicenseLauncher Victoria Launcher Voxscribe - Offline Voice Input Water Sort WaveUp Wikipedia Wristotle Companion Xime IME Yubico Authenticator Some interesting observations brandonp2412 Codeberg The &ldquo;Don&rsquo;t tread on me&rdquo; apps Statistics & Conclusion I love F-Droid. I love what F-Droid stands for and I like the freedom that it gives its users. As a FOSS app maintainer I also have nothing but good things to say about the people behind the project. Maybe they&rsquo;re even a bit too nice, considering how many times repro has failed due to me forgetting to commit before building 😅. But my god is it hard to find human-written software now. Intro # The problem # Sometimes I open F-Droid just to browse. You know, maybe I&rsquo;ll find an app that solves a problem I didn&rsquo;t know I had, or maybe I&rsquo;ll find a better alternative to something I already use. And one day, while browsing, I noticed an app with an obvious and ugly AI generated icon. It&rsquo;s not on the list and I will not shame it, but it got me thinking. How much of F-Droid is AI? As someone who does programming for fun and is only a student, untarnished by honest work, I get my knowledge of what the coding world is like mostly from clickbaity YT videos and Reddit posts of CS professionals. They either describe LLMs as god reincarnate or as glorified autocomplete. Both are, obviously , wrong, but that&rsquo;s not really helpful in determining what I want to know. What is the actual state of programming nowadays? Whenever I use FOSS software, how likely it is that it has been vibe-coded by a rando in an afternoon? How to know if an app is vibe-coded? # That&rsquo;s the trick—you can&rsquo;t. Text just doesn&rsquo;t carry enough meta information for any kind of assessments to be even close to accurate. However, just as that em-dash I used in the first sentence probably triggered an alarm in your brain, there are signs. While that does make the task at hand sound fickle and dependent on happenstance, the signs, especially concerning LLM code repositories, are never too hard to find. You see, the main allure of LLMs is that they allow the developer to be more lazy. That&rsquo;s kind of the whole point! You just prompt, sit back and relax. So it should not surprise you to hear that this attitude is then reflected in everything the vibe-coder touches. Why write a README from scratch? Lol, just let the LLM do it. Do code-review? Naw, just let the LLM review its own changes and then also give it access to the repo so you don&rsquo;t even need to press the commit button. If a project was concerned with looking legitimate, it would be trivial to do such things by hand. But it&rsquo;s low effort all the way down. My biases # This part can be skipped if you don&rsquo;t care about my stance on LLMs, but I need to make my position clear to avoid contributing to the circlejerk too much. To begin with, I&rsquo;d like to acknowledge that LLMs are incredibly useful and capable. As 2026 has progressed, this has become more and more visible, but people being able to one-shot medium scale games and software in half an hour is ridiculously impressive, even if the end product is usually not very good. I also am very much in favor of software getting faster and more secure. The story of the Linux kernel development has shown that LLMs are capable of finding and sometimes even solving many types of code issues. With the pleasantries out of the way though, I have to admit I really hate LLMs and what they have done to programming, related engineering fields, and society as a whole. Their mere existence makes educating yourself and going on fun side projects much less rewarding. Like yeah , I did something, but with an LLM I could have done this in a quarter of the time. And when you do take the black pill and vibe-code, it&rsquo;s even worse. It&rsquo;s not like you did anything. The machine did that. Then there&rsquo;s the atrophying effects on human brains, their unfathomable capability for serving plausibly sounding misinformation and the climate disaster that we&rsquo;re just kinda ignoring. All very fun things to think about. Criteria / The experiment # As mentioned before, there&rsquo;s no way to effectively detect slop, so I propose a rough 3 tier system based on the aesthetics of the repo: Mostly AI This is mostly for projects that have significant LLM smells and means I expect >50% of the code is LLM authored. Any kind of agentic infrastructure automatically lands an app in this tier as I do not believe it is possible to use AI responsibly from within a coding harness. Hard to say / Mostly human / Other Occasional LLM commits either by maintainers or contributors, but mostly looks human. May have an LLM policy which permits certain uses. This tag means I expect <50% of the code is LLM authored. No signs of AI Could not find anything suspicious/Has a strict LLM policy. Note that my rating will still be quite superficial and the tiers quite loose. I did not build a &ldquo;slop detector&rdquo; or anything like that, so my decisions are based on looking at recent commits and their content along with the project&rsquo;s branding. Also note that since I have no way of knowing for sure, there may be errors. I still believe most of my findings to be accurate however. I will also not look at the history of the app. If it has existed since 2014 but recent commits are LLM authored it will be categorized as &ldquo;mostly AI&rdquo; The apps chosen were the batch of updates pushed to F-Droid on September 12, 2026. That amounted to 102 apps, which was quite a lot of work for me to go through.😅 If you just want the results, feel free to skip to the results Apps # Amber # Description: Nostr event signer for Android Repository: https://github.com/greenart7c3/Amber Rating: Mostly AI Justification: All recent commits were done with LLM, PRs accepted from agents, Claude Code and Codex infrastructure present. Aria for Misskey # Description: Dive into the interplanetary microblogging platform 🚀 Repository: https://github.com/poppingmoon/aria Rating: No signs of AI Justification: This is a hard one as commit naming and structure felt a bit suspicious, but I did not see anything else and decided to err on the side of caution Atmo Engine # Description: Animated wallpapers with Atmosphere, Glass, Canvas Sketch and playlists. Repository: https://github.com/saad-khan-rind/nosatmosphereeffect Rating: No signs of AI Justification: Hea