메뉴
BL
MIT Tech Review • 45일 전

차세대 LLM 핵심 기술과 AI 학계의 변화

IMP
8/10
핵심 요약

구글의 트랜스포머(Transformer) 도입 9년 만에 한계에 부딪힌 현재 LLM 구조의 대안이 되는 4가지 혁신 기술을 점검합니다. 또한 민간 기업의 약진으로 연구 환경이 변화하는 가운데 AI 학계가 맞이한 새로운 현실과 주요 기술·정치·비즈니스 이슈를 총망라했습니다.

번역된 본문

이 기사는 매일 기술 세계의 최신 소식을 전하는 평일용 뉴스레터인 'The Download'의 오늘자 버전입니다.

이 스타트업들이 노리는 차세대 대형 언어 모델(LLM)의 패러다임 구글 연구진이 트랜스포머(Transformer)를 발표한 지 9년이 지난 지금, 이 신경망 계열 모델은 모든 주요 대형 언어 모델(LLM) 내부를 구동하는 핵심 엔진이 되었습니다. 하지만 트랜스포머도 이제는 노후화의 조짐을 보이기 시작했습니다. LLM의 규모가 커지고 성능이 향상되면서 트랜스포머는 오히려 병목 현상의 원인이 되고 있습니다. 처리해야 할 텍스트 양이 늘어날수록 이들의 밀집 어텐션(attention) 메커니즘은 비용이 기하급수적으로 증가하며, 방대한 정보를 한 번에 처리하고 추적하는 데에는 한계를 드러내고 있습니다. 여기에는 이러한 트랜스포머의 문제를 해결하기 위한 4가지 새로운 아이디어가 있습니다. 이 혁신들은 LLM을 영원히 바꿔놓아 더 빠르고, 훨씬 효율적이며, 어쩌면 더 똑똑하게 만들 수 있습니다. — Will Douglas Heaven 이 기사는 업계, 트렌드 및 기술을 분석하여 미래를 먼저 보여주는 MIT Technology Review의 'What's Next' 시리즈에서 가져왔습니다. 나머지 기사는 이곳에서 확인할 수 있습니다.

AI 교수들이 현재 학계의 새로운 현실과 타협하고 있습니다 — Grace Huckins 지난주 저는 캘리포니아 마운틴뷰의 한 호텔에 방문하여, 세계에서 가장 훌륭하고 유망한 AI 연구자들과 함께했습니다. 에릭 슈미트(Eric Schmidt)와 웬디 슈미트(Wendy Schmidt)의 자금 지원을 받아 AI 관련 연구를 지원하는 'Schmidt Sciences AI2050' 프로그램 행사에서, 저는 원탁 인터뷰를 진행하고 미디어 트레이닝 연설을 했습니다. 이 프로그램의 펠로우 명단은 AI 업계의 쟁쟁한 인사들로 구성되어 있으며, 모두가 참석하지는 않았지만 제가 모퉁이를 돌 때마다 과거 인터뷰했거나 그 연구에 감탄했던 과학자들과 마주쳤습니다. 현재 AI2050 그룹의 대부분을 차지하는 대학 기반 AI 연구자들에게는 다소 당혹스러운 시기입니다. 그 이유가 무엇인지, 그리고 앞으로 무슨 일이 일어날지는 Grace의 기사를 통해 확인해 보시기 바랍니다. 이 기사는 매주 월요일 받아보는 AI 주간 뉴스레터인 The Algorithm에서 가져왔습니다. 수신함에서 받아보려면 구독하세요.

오늘의 필수 뉴스 오늘날 기술과 관련하여 가장 흥미롭고, 중요하고, 무섭고, 매혹적인 기사를 인터넷에서 찾아 정리했습니다.

  1. 엔비디아, 월스트리트로부터 5천억 달러의 AI 인프라 투자 유치 블랙록(BlackRock), 골드만 삭스(Goldman Sachs) 등 4개 기관과 투자 계약을 체결했습니다. (BBC) + 기관 투자자들 사이에서 AI 컴퓨팅에 대한 매력을 보여줍니다. (로이터 $) + 이에 따라 AI 인프라는 새로운 자산 등급으로 자리 잡고 있습니다. (CNBC)
  2. 마크 저커버그의 새로운 비전: 오픈소스 AI가 미국을 구할 수 있다 맞춤형 '초지능(Superintelligence)'에 대한 유토피아적인 비전을 제시했습니다. (가디언) + 이는 메타(Meta)의 새로운 오픈소스 모델이 발표된 날과 같습니다. (NYT $) + 저커버그는 더 많은 모델을 출시할 계획이라고 밝혔습니다. (WSJ $) + 또한 중국의 오픈 웨이트(Open-weight) 개발자들과 메타가 경쟁할 것이라고 덧붙였습니다. (SCMP)
  3. 버니 샌더스, 실리콘밸리에 'AI 개발 중단' 촉구 AI 거대 기업들은 안전을 위해 필요한 경우 스스로 개발을 중단하겠다고 약속했다고 지적했습니다. + 또 아무런 조치가 취해지지 않으면 입법자들이 개입할 것이라고 경고했습니다. (가디언) + 하원 민주당원들은 이미 통제 불능 모델에 대해 AI 리더들을 압박하고 있습니다. (WP $) + AI에 대한 대중적인 반발이 거세지고 있습니다. (MIT Technology Review)
  4. 미국 법원, 수천 건의 소셜 미디어 소송 진행 허용 이 소송들은 메타, 틱톡, 구글, 스냅챗이 사용하는 중독성 유도 시스템을 겨냥하고 있습니다. (Axios) + 이들이 젊은 사용자를 고객으로 묶어두기 위해 플랫폼을 설계했다고 주장합니다. (로이터 $) + 우리는 인터넷을 고칠 수 있을까요? (MIT Technology Review)
  5. 유니트리(Unitree) 기업공개(IPO), 개인 투자자들로부터 8천 배 이상의 청약 경쟁률 기록 이 중국 휴머노이드 로봇 기업은 상장에 앞서 9억 달러를 조달했습니다. (로이터 $) + 상하이 IPO 가격 책정으로 이 회사의 가치는 90억 달러로 평가되었습니다. (FT $)
  6. 차량 추적 카메라 업체 Flock, 양당의 비판에 직면하다 이 감시 네트워크는 미국 전역에 빠르게 확산했습니다. (NYT $) + Flock은 드론을 이용해 소매치기를 추적할 계획도 세우고 있습니다. (MIT Technology Review)
  7. 중국, AI 관계 규제 착수 베이징 당국은 감정적 상호작용이 가능한 AI에 대한 새로운 규정을 도입했습니다. (Rest of World) + 사람들이 챗봇과 사랑에 빠지는 것은 놀라울 정도로 쉽습니다.
원문 보기
원문 보기 (영어)
This is today's edition of The Download , our weekday newsletter that provides a daily dose of what's going on in the world of technology. These startups are chasing the next big thing in LLMs Nine years after Google researchers introduced the transformer, this family of neural networks has become the engine inside every major large language model. But transformers are starting to show their age. As LLMs get bigger and better, transformers have become a bottleneck. Their dense attention mechanism becomes increasingly expensive as the amount of text grows, and they’re not great at keeping track of a lot of information at once. Here are four new ideas for how to solve the transformer problem —innovations that could change LLMs for good, making them faster, far more efficient, and (maybe) even smarter. —Will Douglas Heaven This story is from MIT Technology Review’s What’s Next series, which looks across industries, trends, and technologies to give you a first look at the future. You can read the rest of them here . AI professors are negotiating the new realities of academic research —Grace Huckins Last week, I headed to a hotel in Mountain View, California, to join some of the most accomplished, and some of the most promising, AI researchers in the world. I was hosting roundtable interviews and speaking at a media training for a convening of the Schmidt Sciences AI2050 program, an initiative funded by Eric and Wendy Schmidt that supports academics whose work involves AI. The fellows list is a who’s who of AI luminaries, and though not all of them made it out to the Bay, every time I turned a corner I saw a scientist whom I’d interviewed previously or whose research I admired. It’s a weird time for university AI researchers, who make up most of the AI2050 group. Read Grace’s story to find out why, and what could be coming next . This story is from The Algorithm, our weekly AI newsletter. Sign up to receive it in your inbox every Monday. The must-reads I’ve combed the internet to find you today’s most fun/important/scary/fascinating stories about technology. 1 Nvidia has secured $500 billion from Wall Street for AI infrastructure It’s struck deals with BlackRock, Goldman Sachs, and four others. ( BBC ) + Showing the pull of AI compute for ⁠institutional investors. ( Reuters $) + And that AI infrastructure is becoming a new asset class. ( CNBC ) 2 Mark Zuckerberg's new manifesto says open-source AI can save the US It presents a utopian vision of personalized “superintelligence.” ( Guardian ) + And arrived the same day as Meta’s new, open-source model. ( NYT $) + Zuckerberg said he plans to launch more of these models. ( WSJ $) + And pit Meta against Chinese open-weight developers. ( SCMP ) 3 Bernie Sanders has called on Silicon Valley to “pause AI development” He noted that AI giants have pledged to do this if necessary for safety. + And warned that lawmakers will step in if no action is taken. ( Guardian ) + House Democrats are already pressing AI leaders over rogue models. ( WP $) + A populist backlash is building against AI. ( MIT Technology Review ) 4 A US court will allow thousands of social media lawsuits to proceed The suits target addictive mechanisms used by Meta, TikTok, Google, and Snapchat. ( Axios ) + They claim the platforms are designed to hook young users. ( Reuters $) + Can we repair the internet? ( MIT Technology Review ) 5 Unitree's IPO is more than 8,000 times oversubscribed by retail The Chinese humanoid firm raised $900 million ahead of its listing. ( Reuters $) + Its pricing for the Shanghai IPO values the company at $9 billion. ( FT $) 6 Flock’s car-tracking cameras are facing a bipartisan backlash The surveillance network has spread rapidly across the US. ( NYT $) + Flock also plans to chase shoplifters with drones. ( MIT Technology Review ) 7 China is breaking up AI relationships Beijing has introduced new rules for emotionally interactive AI. ( Rest of World ) + It’s surprisingly easy to fall for a chatbot. ( MIT Technology Review ) 8 An AI tool claims to pick the best 1% of scientific papers But researchers doubt that AI can reliably judge scientific quality. ( Nature ) 9 The AI slop backlash is working It’s pushing platforms to restrict AI-generated content. ( Wired $) 10 An 82-year-old rejected $26 million to turn her farm into a data center She criticised the environmental impacts of data centers. ( Fortune ) Quote of the day “It is not too late to avoid disaster. Stop building machines that humans cannot control.” —Senator Bernie Sanders urges Sam Altman, Dario Amodei, and Mark Zuckerberg to pause all AI development in a letter . One More Thing The race to make the perfect baby is creating an ethical mess A new field of science is using genetic sequencing to predict what kind of person an embryo might become. Some parents turn to these tests to avoid devastating genetic disorders, while a much smaller group are willing to pay tens of thousands of dollars to optimize for intelligence, appearance, and personality. Customers, however, may not be getting what they’re paying for. Genetics experts have highlighted the potential deficiencies of this testing for years, while its underlying assumptions have made these companies a political lightning rod. As this technology edges toward the mainstream, scientists and ethicists are racing to confront the implications—for our social contract, for future generations, and for our very understanding of what it means to be human. Read the full story . —Julia Black We can still have nice things A place for comfort, fun, and distraction to brighten up your day. (Got any ideas? Drop me a line .) + An eagle-eyed border collie is taking the game of fetch into new waters . + musicForprogramming has made a valiant attempt to produce the perfect tunes for sustained concentration. + When kids design playgrounds , they create a cheerful mix of giant chess, pink basketball courts—and lava. + This power metal version of the “Back to the Future” music is an epic reinvention of the film’s classic theme. Deep Dive The Download The Download: Claude’s inner workings and OpenAI’s “super app” Plus: OpenAI has unveiled its long-awaited "super app." By Thomas Macaulay archive page The Download: Claude’s inner workings, and the future of world models Plus: New York has become the first state to enact a data center moratorium. By Thomas Macaulay archive page The Download: the future of chipmaking and Anthropic’s government clash Plus: Meta is pausing an AI training program that tracks workers’ keystrokes. By Thomas Macaulay archive page The Download: a reality check for geoengineering and the science of interoception Plus: SpaceX is now valued higher than Amazon. By Thomas Macaulay archive page Stay connected Illustration by Rose Wong Get the latest updates from MIT Technology Review Discover special offers, top stories, upcoming events, and more. Enter your email Privacy Policy Thank you for submitting your email! Explore more newsletters It looks like something went wrong. We’re having trouble saving your preferences. Try refreshing this page and updating them one more time. If you continue to get this message, reach out to us at customer-service@technologyreview.com with a list of newsletters you’d like to receive.