Qwen 3.8 27B, Cerebras에서 초당 1,500토큰으로 제공
Cerebras가 Qwen 3.8 27B 모델을 초당 1,500토큰으로 제공하기 시작했습니다.
Cerebras introduces Qwen 3.8 27B model offering 1,500 tokens per second.
AI가 선별한 아티클
Cerebras가 Qwen 3.8 27B 모델을 초당 1,500토큰으로 제공하기 시작했습니다.
Cerebras introduces Qwen 3.8 27B model offering 1,500 tokens per second.
Cohere가 복잡한 문서에서 구조화된 데이터를 추출하는 Parse 5를 출시했습니다.
Cohere has launched Parse 5, a multimodal model for extracting structured data from complex documents.
알리바바가 Qwen3.8-Flash를 출시하며 Qwen4 아키텍처의 초기 미리보기를 공개했습니다.
Alibaba has released Qwen3.8-Flash, an early preview of the Qwen4 architecture.
Z.ai가 GLM-5.3-Flash를 출시하며 비용 효율성을 강조했습니다.
Z.ai launched GLM-5.3-Flash, highlighting its cost efficiency.
Qwen 3.8-Flash-Next 모델이 2026년 8월 26일 공개될 예정이다.
The Qwen 3.8-Flash-Next model is set to be released on August 26, 2026.
3B 오픈 가중치 모델 'Shieldstral'이 멀티모달 콘텐츠 검열을 위한 안전 분류기로 공개되었다.
'Shieldstral', a 3B open-weight model for multimodal content censorship, has been released.
Mistral이 3B 오픈 가중치 모델 'Shieldstral'을 발표했습니다.
Mistral announces 'Shieldstral', a 3B open-weights model for multimodal moderation.
알리바바가 Qwen3.8-Max 모델을 출시했다.
Alibaba has launched the Qwen3.8-Max model.
알리바바의 AI가 16일간 코드 작성을 했으며, 모든 커밋이 GitHub에 기록됨.
Alibaba's AI wrote code for 16 days straight, with every commit available on GitHub.
Kakao의 Kanana 팀이 사용자 맞춤형 AI 모델을 개발하는 과정에 대해 소개합니다.
Kakao's Kanana team introduces the process of developing user-oriented AI models.
카카오는 Kanana-o 음성 생성 모델의 고도화 과정을 소개합니다.
Kakao introduces the enhancement process of the Kanana-o speech generation model.
Hugging Face에 Kimi-K3 모델 공개, 멀티모달 작업 지원.
Hugging Face releases Kimi-K3 model, supporting multimodal tasks.
Claude 애플리케이션 구현을 위한 실전 가이드와 예제를 제공하는 문서입니다.
This document provides practical guides and examples for implementing Claude applications.
FLUX 3 모델이 멀티모달 학습을 통해 이미지, 비디오, 오디오를 통합적으로 다룬다.
FLUX 3 model integrates image, video, and audio through multimodal learning.
구글이 새로운 제미나이 AI 모델을 출시했습니다.
Google has launched new Gemini AI models.
Inkling은 975B 파라미터의 오픈 웨이트 모델로 다양한 미디어 입력을 처리합니다.
Inkling is an open weight model with 975B parameters that supports various media inputs.
Muse Spark 1.1은 멀티모달 추론 능력을 향상시키며 에이전트 작업에 초점을 맞춘 모델이다.
Muse Spark 1.1 enhances multimodal reasoning for agent tasks.
ECCV 2026에서 열리는 MARS2 워크숍에 대한 논의가 이루어지고 있다.
Discussion is underway about the MARS2 Workshop at ECCV 2026.
구글의 Gemini-3-Flash 모델에 대한 초보자 가이드.
A beginner's guide to Google's Gemini-3-Flash AI model.
BugCapture로 AI-ready 버그 보고서를 자동 생성하는 방법을 소개합니다.
Introducing BugCapture to automate AI-ready bug reports from screen recordings.
KOLongDoc 벤치마크는 한국어 긴 문서를 읽는 VLM의 성능을 평가합니다.
KOLongDoc benchmark evaluates VLM performance on Korean long documents.
Gemma 4 12B는 멀티모달 지능을 위한 인코더 없는 모델이다.
Gemma 4 12B is an encoder-free model for multimodal intelligence.
Gemma 4는 Raspberry Pi에서 연구 워크스테이션까지 가는 네 가지 멀티모달 모델입니다.
Gemma 4 consists of four multimodal models ranging from Raspberry Pi to research workstation.
OpenAI의 GPT-4는 이미지와 텍스트 입력을 처리하는 대형 다중모달 모델이다.
OpenAI's GPT-4 is a large multimodal model that processes image and text inputs.