7·ai-ml·기타·InfoQ·2026. 08. 29. FreeToken은 소비자 하드웨어에서 Mixture-of-Experts 모델을 최적화하는 오픈소스 추론 엔진이다.
FreeToken is an open-source inference engine that optimizes Mixture-of-Experts models on consumer hardware.
8·ai-ml·릴리즈·The New Stack·2026. 06. 04. Nvidia의 5500억 매개변수를 가진 모델이 출시됐다.
Nvidia releases its 550-billion-parameter model Nemotron 3 Ultra.
8·ai-ml·분석·r/MachineLearning·2026. 05. 18. Residual Coupling을 이용한 LLM의 수평 확장 및 성능 개선 방법.
Method for horizontally scaling LLMs using Residual Coupling for improved performance.
6·ai-ml·분석·Dev.to·2026. 05. 16. Gemma 4 모델을 테스트하여 GPT-4o-mini와의 성능 비교를 진행한 결과, 서로 다른 아키텍처에서 상반된 반응을 보였다.
I tested Gemma 4 variants against GPT-4o-mini and found differing responses based on architecture.
모든 아티클을 불러왔습니다.