DeepSeek

Latest AI news, models and releases from DeepSeek. ['DeepSeek', 'DeepSeek 3.1', 'DeepSeek-4-pro', 'DeepSeek API', 'deepseek-coder-v2', 'DeepSeek-Coder-V2-Lite', 'DeepSeek-GRM', 'DeepSeek large model', 'DeepSeekMath-V2', 'DeepSeekMoE', 'DeepSeek-Prover-V1.5-Base', 'DeepSeek-Prover-V2', 'DeepSeek-Prover-V2-671B', 'DeepSeek-Prover-V2-7B', 'DeepSeek R1', 'DeepSeek-R1', 'DeepSeek R1-0528', 'DeepSeek-R1-0528', 'DeepSeek-R1-671B', 'DeepSeek-R1-Distilled-Llama-70B', 'DeepSeek-R1-Distilled-Qwen-32B', 'DeepSeek-R1-Distill-Qwen-32B', 'DeepSeek-R1-Zero', 'DeepSeek R2', 'DeepSeek-R2', 'DeepSeek Sparse Attention (DSA)', 'DeepSeek V1', 'DeepSeek V2', 'DeepSeek-V2', 'DeepSeek-V2.5', 'DeepSeek V3', 'DeepSeek-V3', 'DeepSeek V3.1', 'DeepSeek V3.1-Terminus', 'DeepSeek v3.2', 'DeepSeek V3.2', 'DeepSeek-V3.2', 'DeepSeek V3.2-Exp', 'DeepSeek-V3-Base', 'DeepSeek V3/R1']

Applications 🇷🇺

Automating Text-to-SQL in WMS: Choosing a Local LLM Without GPU

A developer at the logistics company Aerosib-C automated the generation of SQL queries for invoicing using LLM models. Due to security requirements, local models were used, tested on a Linux server with 30 GB of RAM and no GPU. The best result was achieved by gemma4:26b, although response times reached tens of minutes. Coder models were outperformed by reasoning models. Plans include building an expert system for service classification.

DeepSeekDeepSeek Google/DeepMindGoogle/DeepMind Alibaba/QwenAlibaba/Qwen
Habr — хаб ИИ24.07 · 03:02
Fresh news