Shieldstral - 멀티모달 콘텐츠 검열을 위한 3B 오픈 가중치 모델
3B 오픈 가중치 모델 'Shieldstral'이 멀티모달 콘텐츠 검열을 위한 안전 분류기로 공개되었다.
'Shieldstral', a 3B open-weight model for multimodal content censorship, has been released.
AI가 선별한 아티클
3B 오픈 가중치 모델 'Shieldstral'이 멀티모달 콘텐츠 검열을 위한 안전 분류기로 공개되었다.
'Shieldstral', a 3B open-weight model for multimodal content censorship, has been released.
Mistral이 3B 오픈 가중치 모델 'Shieldstral'을 발표했습니다.
Mistral announces 'Shieldstral', a 3B open-weights model for multimodal moderation.
알리바바가 Qwen3.8-Max 모델을 출시했다.
Alibaba has launched the Qwen3.8-Max model.
알리바바의 AI가 16일간 코드 작성을 했으며, 모든 커밋이 GitHub에 기록됨.
Alibaba's AI wrote code for 16 days straight, with every commit available on GitHub.
Kakao의 Kanana 팀이 사용자 맞춤형 AI 모델을 개발하는 과정에 대해 소개합니다.
Kakao's Kanana team introduces the process of developing user-oriented AI models.
카카오는 Kanana-o 음성 생성 모델의 고도화 과정을 소개합니다.
Kakao introduces the enhancement process of the Kanana-o speech generation model.
Hugging Face에 Kimi-K3 모델 공개, 멀티모달 작업 지원.
Hugging Face releases Kimi-K3 model, supporting multimodal tasks.
Claude 애플리케이션 구현을 위한 실전 가이드와 예제를 제공하는 문서입니다.
This document provides practical guides and examples for implementing Claude applications.
FLUX 3 모델이 멀티모달 학습을 통해 이미지, 비디오, 오디오를 통합적으로 다룬다.
FLUX 3 model integrates image, video, and audio through multimodal learning.
구글이 새로운 제미나이 AI 모델을 출시했습니다.
Google has launched new Gemini AI models.
Inkling은 975B 파라미터의 오픈 웨이트 모델로 다양한 미디어 입력을 처리합니다.
Inkling is an open weight model with 975B parameters that supports various media inputs.
Muse Spark 1.1은 멀티모달 추론 능력을 향상시키며 에이전트 작업에 초점을 맞춘 모델이다.
Muse Spark 1.1 enhances multimodal reasoning for agent tasks.
ECCV 2026에서 열리는 MARS2 워크숍에 대한 논의가 이루어지고 있다.
Discussion is underway about the MARS2 Workshop at ECCV 2026.
구글의 Gemini-3-Flash 모델에 대한 초보자 가이드.
A beginner's guide to Google's Gemini-3-Flash AI model.
BugCapture로 AI-ready 버그 보고서를 자동 생성하는 방법을 소개합니다.
Introducing BugCapture to automate AI-ready bug reports from screen recordings.
KOLongDoc 벤치마크는 한국어 긴 문서를 읽는 VLM의 성능을 평가합니다.
KOLongDoc benchmark evaluates VLM performance on Korean long documents.
Gemma 4 12B는 멀티모달 지능을 위한 인코더 없는 모델이다.
Gemma 4 12B is an encoder-free model for multimodal intelligence.
Gemma 4는 Raspberry Pi에서 연구 워크스테이션까지 가는 네 가지 멀티모달 모델입니다.
Gemma 4 consists of four multimodal models ranging from Raspberry Pi to research workstation.
OpenAI의 GPT-4는 이미지와 텍스트 입력을 처리하는 대형 다중모달 모델이다.
OpenAI's GPT-4 is a large multimodal model that processes image and text inputs.