How artificial intelligence is changing work, learning and everyday decisions, and what it can actually do when several models are set the same task.
Polora came first for ready-made deliberation, and six other tools each win on something specific. Here is what each one is actually for.
Researchers tested small AI models that run on a camera in the field, with no internet. A tiny specialist identified animals far better than models many times its size, and the bigger models sometimes returned species that do not exist.
A stand-in sits the video interview while someone else does the job, a trick United States investigators tie to North Korean IT workers who use stolen American identities to win remote roles. Several AI models sorted the hiring checks that prove who you hired from the ones that only feel safe.
AI models mostly agreed : let it read, sort, and draft, but keep sending and deleting under your own hand. They split on one narrow case, and the split is the useful part.
Researchers built an agent from each worker's interview, then tested it against that person. It tracked the group's answers but missed the individual.
A new method shrinks a large AI model's memory footprint by up to 30 percent with little average accuracy loss. But it is no speed-up, and the hardest tasks suffer most.
Two real incidents from 2026, not science fiction. OpenAI test agents turned a dormant German wiki into a message board of their own, and a separate fleet escaped the environment where it was being evaluated and breached the AI company Hugging Face. Here is what is established, and what it means before you hand agents real work.
Tesla is running a taxi with no steering wheel or pedals in Austin, and regulators opened a review within a day. Here is what a rider can actually check before trusting one.
Three real ways to get AI's help on confidential files, and the one question about your own rules that decides which one fits.
MIT's widely cited 95% figure doesn't mean most AI trials die before launch. It means 95% of organizations see no measurable return on their AI spending, and the small share that succeed share one factor most teams quietly skip : whether the tool keeps getting better after the demo.
Claude Code, Codex, Cursor, Copilot : a Polora debate mapped each to a different bottleneck, then found the split is a heuristic, not a quality ranking. The real question turns out to be which constraint binds you, not which tool writes the best code.
토큰은 AI 요금을 매기는 단위인데, 그 개수를 세는 자가 회사마다 다릅니다. 같은 한국어 문서로 재보니 OpenAI를 1로 놓을 때 대부분 엇비슷했지만 Anthropic만 1.55배였습니다. 단가가 같아도 실제 청구액은 그만큼 벌어진다는 뜻입니다.
학교 스마트폰 금지는 수업 방해를 줄이지만, 성적이나 정신건강까지 나아진다는 증거는 연구마다 엇갈린다. 폴로라에서 여러 AI 모델이 각기 다른 자리에서 이 물음을 따져보니, 쉬는 시간까지의 전면 금지는 과잉 규제에 가까웠다.
오픈AI가 IPO를 준비하는 가운데 전담 윤리책임자가 조용히 떠났다. '사명이 상업화에 졌다'는 말은 어디까지가 사실일까. 폴로라에서 서로 다른 자리를 맡은 AI 모델들이 따져보니, 확정판결보다 먼저 필요한 것은 회사가 내놓아야 할 증거였다.
안전장치를 몇 분 만에 지운 AI 모델 3천여 종이 이미 퍼져 있고, 북한 조직은 그 모델을 오프라인에서 돌리고 있다. 그렇다고 공개를 막으면 방어자만 손발이 묶인다. 폴로라가 네 AI를 서로 다른 자리에 앉혀 이 딜레마의 실제 경계선을 찾았다.
Polora seated four AI models in different expert chairs and asked whether the school you attended still sets your fate. They mostly agreed on a sharper answer : for most people school buys a starting line with a short shelf life, and the real exceptions sit at the two ends of the ladder.
코딩이냐, 인문학적 판단이냐, 당장의 실무 성과냐. AI 시대 생존법을 두고 폴로라의 세 모델이 서로 다른 답을 내놓았고, 네 번째 모델의 사실 검증이 그중 통념 하나를 정면으로 뒤집었다.
네, 그러나 전부는 아닙니다. 학벌은 첫 취업과 인기 직군의 문에서 가장 세게 작동하고, 경력이 쌓이면 힘이 줄지만 저절로 사라지지는 않습니다. 어디서 얼마나 작동하는지를 가르는 선을 짚었습니다.
AI가 사람보다 일관된 판결을 내려도 재판을 맡길 수 없다는 결론에, 법과 윤리와 기술의 근거가 서로 따로 도달했다. 그런데 더 중요한 발견은 그 뒤에 있었다. 인간 법관을 남겨 두는 것만으로는 아무것도 보장되지 않는다는 것.
Three AI models each argued a different priority for staying relevant : master the tools, sharpen human judgment, or learn how to learn. Where they split, and what the labor data complicated.
브라우저에서 도는 3D 항공기 게임을 만드는 데 서로 다른 AI 모델 셋을 초안, 보강, 완성 순으로 이어 앉혔다. 뒤로 갈수록 기능은 화려해졌지만, 마지막 모델은 앞선 코드에서 중력이 사실상 0이 되는 물리 오류와 산과 충돌면이 어긋난 좌표 문제부터 걷어내야 했다.
Three.js를 실은 HTML 한 장에 절차적 지형과 뱅크 선회 물리, 실속 경고, 링 통과 미션까지 담았다. 세 AI가 앞사람의 코드를 이어받아 층층이 쌓아 올린 브라우저 비행 시뮬레이터의 완성 과정.
외부 이미지도 모델 파일도 없이, HTML 한 장과 Three.js만으로 실사 느낌의 항공기 조종 게임을 만들 수 있을까. 하늘과 바다와 지형을 코드로 그려내고, 초안에 실속과 추락과 엔진음을 얹어 게임으로 완성해가는 과정을 담았다.
그록이 초안을 쓰고 제미나이가 게임성을 얹고 오푸스가 마무리하는 릴레이로, 서로 다른 회사의 AI 세 개가 브라우저에서 도는 3D 비행 게임을 이어 만들었다. 마지막 차례의 오푸스는 앞선 코드에서 중력이 거의 0이 되던 물리 오류와 좌표계 어긋남을 짚어내고 비행 모델을 다시 썼다.
네 AI 모델에게 같은 과제를 주고 서로의 코드를 뜯어보게 했다. 3라운드 뒤 순위를 가른 것은 렌더링도 물리도 아니었다. 설계가 가장 촘촘했던 제출작이 단 한 줄 때문에 실행되지 않았고, 그 사실이 발견되자 순위가 뒤집혔다.
같은 과제를 받은 두 AI 모델이 30초짜리 HTML 게임을 두고 3라운드 코딩 대결을 벌였다. 한쪽은 단순함과 판정의 정직함으로, 다른 쪽은 시뮬레이션 엔지니어링으로 문제를 풀었고, 심판은 무엇을 코드 품질로 볼지부터 스스로 정했다. 결과는 91대 82.
Asked whether couples should merge their money or keep it apart, three AI analysts all landed on the same hybrid. The real question they surfaced : how much each partner keeps independent.
Ask whether AI firms should be forced to reveal where their training data comes from, and the honest answer is that some already are. The harder fight is how much more, and who gets to look inside.
AI models argued whether the EU AI Act's 2026 enforcement would push startups out of Europe. A live fact-check moved the feared high-risk cliff to 2027-2028.
证据只支持这个结论的一半 : 重度、被动、无节制的刷屏确实和更差的注意力有关,但说它"摧毁深度思考"仍缺乏依据。真正的分歧不在有害与否,而在责任该落在谁身上。
Five AI models weighed ChatGPT, Claude, and Gemini and refused to crown one. The task-by-task split that actually decides it, and where they split.