메뉴
HN
Hacker News 20일 전

LLM 사용에 지쳐버린 개발자의 고백

IMP
5/10
핵심 요약

한 개발자가 매일같이 코딩과 검색에 LLM을 사용하면서 쌓인 'LLM 번아웃'을 토로합니다. LLM이 보여주는 기계적인 문체와 반복되는 환각, 잘못된 가정 등을 끊임없이 검토하고 수정하는 과정이 큰 피로감으로 다가오고 있다고 설명합니다. AI 도구의 생산성은 인정하면서도, 동일한 패턴의 오류와 글쓰기 스타일에 반복적으로 노출되는 것에 대한 개발자들의 공감대를 이끌어내는 글입니다.

번역된 본문

나는 LLM을 매우 자주 사용한다. 현직 개발자들의 기준으로 볼 때 내 사용량은 평균적이고, 사용 방식은 꽤 단순한 편일 것이다. 나는 한 번에 하나의 작업만 진행하며, (회사에서는) Claude Code나 (지금 당장은 집에서) Codex와 함께 작업 내용을 논의한다. 때로는 AI 어시스턴트가 코드를 작성하게 두지만, 나는 그 결과물을 꼼꼼히 읽고 이해한 뒤 수정한다. 나를 자율형 에이전트나 에이전트 오케스트레이션의 깊은 물에 빠뜨리지는 않는다. 그럼에도 나는 매일 업무와 개인 시간을 통틀어 몇 시간씩 LLM과 상호작용하는 데 보낸다. 이건 몇 년 전보다 훨씬 많아진 양이며, 아마 AI가 생성한 텍스트를 읽지 않는 날은 하루도 없을 것이다.

내 업무는 코드를 직접 디자인하고 작성하는 것에서, 코드를 디자인하고 그 설계를 LLM에게 설명해준 뒤, LLM이 작성한 코드를 리뷰하고 마지막에야 코드를 다듬는 방식으로 바뀌었다. LLM과 작업하는 과정은 내가 미처 생각하지 못했거나 몰랐던 접근 방식을 접할 수 있게 해준다. 또한 내가 깊이 알지 못하는 분야에서도 훨씬 편안함을 느끼게 해줬다. 지금 내 주요 프로젝트는 우리 코드베이스에 대규모 무인 감독(unsupervised) 코드 생성을 위한 프레임워크를 구축하는 것이다. Claude와 함께 도구를 만들지 않을 때는, 무인 에이전트(Qwen)의 결과물을 샅샅이 뒤지고 있다. 어느 쪽이든 나는 쉬지 않고 LLM이 생성한 콘텐츠를 읽어야 한다.

무언가를 알고 싶을 때, 내가 직접 확인해봐야 할 사이트가 정확히 있는 게 아니라면 아마 ChatGPT에게 물어보거나 Gemini의 요약본을 읽을 것이다. LLM의 답변이 틀렸을 때는 어쩔 수 없이 다시 웹 서핑으로 돌아가야 하지만, 특히 쓸모없는 AI 생성 글들이 검색 결과를 도배하고 있을 때는 대다수의 일상적인 질문에는 LLM 정도면 충분히 만족스럽다. 이런 생활이 벌써 1년째이며, 앞으로 멈출 생각은 없다. LLM을 쓸 때 생산성이 훨씬 높아진다고 느끼며, 이를 효과적으로 사용하는 법을 계속 배우는 것은 가치 있는 일이라고 생각한다.

하지만 지난 몇 달간 내 마음가짐이 조금 변했다. 내 안의 작은 부분 중 하나는 이제 LLM의 결과물을 읽는 것을 두려워하기 시작했다. 어떤 내용이 나올지 이미 알고 있기 때문이다. 잘못된 가정과 환각(Hallucination). 강조조, 그리고 기계적으로 끊어지는 문장 조각들. ✨ 과도한 이모티콘들 🚀. 이건 나만의 문제가 아니다. 이들은 명백한 패턴이다(🤮). 따로 떼어놓고 보면 이런 짜증 나는 요소 중 어느 것도 나를 크게 흔들지 않는다. 하지만 이 모든 게 합쳐지면, 나는 금세 LLM의 글에 진절머리가 난다.

나는 LLM을 비난하려는 게 아니다. 인간도 실수를 한다. 우리 역시 똑같이 믿을 수 없거나 짜증 날 수 있다. 문제는 '반복'이다. LLM은 늘 같은 스타일로 글을 쓰고, 똑같은 종류의 실수를 저지른다. 똑같은 상황을 끊임없이 마주하는 것이 나를 지치게 만든다. 인터페이스에서 개인화 기능을 제공한다면 사용할 수 있겠지만, 일부 특유의 기괴한 습성들은 결국 스며나온다. 그리고 당연하게도, 내가 다른 사람들이 생성한 콘텐츠의 스타일까지 통제할 수는 없다.

아직 이 느낌에 어떻게 대처해야 할지 모르겠다. 내가 이토록 신경 쓰일 줄은 몰랐다. 불안정한 도구에 대한 좌절감은 이해가 가지만, 글쓰기 패턴 마저 내 속을 뒤집어 놓는다. 당분간은 이를 악물고 버텨야겠다. 속이 다 뒤집어지지 않기를 바라면서.

원문 보기
원문 보기 (영어)
I use LLMs a lot. By current dev standards, my usage rate is probably average, and my methods are probably primitive. I work on one task at a time and discuss it with Claude Code (at work) or Codex (at home, for now). Sometimes, I let the assistant write code, but I read the output thoroughly, understand it, and revise it. I’m not in the deep end of autonomous agents or agent orchestration. Still, I spend hours each day interacting with LLMs across work and home. That’s a hell of a lot more than I did a few years ago, and I probably don’t go a day without reading AI-generated text. My job has changed from designing and writing code to designing code, describing the design to an LLM, reviewing code the LLM produces, and then finally writing code. The LLM steps expose me to approaches I might not have considered or been aware of. I also feel more comfortable in areas where I don’t have deep knowledge. My main project right now is to establish a framework for large-scale, unsupervised code generation in our codebase. When I’m not working with Claude to create tooling, I’m sifting through the unsupervised agent’s (Qwen’s) output. Either way, I’m reading LLM content. If I want to know something, I’ll probably ask ChatGPT or read Gemini’s overview unless I know what sites I want to check. I still have to fall back to browsing when the LLM’s answer is wrong, but it’s good enough for many casual queries, especially when useless AI-generated articles clutter the search results. It’s been this way for about a year, and I don’t see myself stopping. I feel more productive with LLMs, and I think continually learning how to use them effectively is valuable. However, my disposition has changed a bit in the last few months. Some small part of me has started to dread reading LLM output because I know what I’m going to find. False assumptions and hallucinations. Emphatic, staccato fragments. ✨ Excessive emojis 🚀. It’s not just me—these are real patterns (🤮). On their own, none of these annoyances gets to me. Together, though, they’ve gotten me sick of LLM writing in a hurry. I’m not trying to condemn LLMs. Humans are fallible, too—we can be just as unreliable or annoying. The problem is repetition. LLMs write in the same style, and they make the same kinds of mistakes. Dealing with the same thing over and over is wearing me out. I can use personalization features if the interface offers them, but some idiosyncrasies seep through. And of course, I don’t control the style of content generated by other people. I don’t know how to deal with this feeling yet. I didn’t expect to be so bothered by it. Frustration at a flaky tool is understandable, but the writing patterns grind my gears, too. For now, I’ll grit my teeth and hope I don’t lose my lunch.