Discover Great Companies Discover Great Clients
Companies Salaries Jobs
Jobs |
Jobs Companies Salary Majors

深圳 运维 jobs (salary & requirements)

深圳市工匠行科技有限公司 运维 salary range: 8K - 10K, where 100% of positions earn ¥8-10K
¥8-10K
100% of positions earn

Note: the average salary is analyzed based on job postings published by the company. We recommend reviewing it together with position type, education, region and experience.

深圳市工匠行科技有限公司 运维 salary changes over the years

Note: the data depends on salary samples of online job postings in the corresponding years and does not fully represent the actual situation within the company. For reference only.

运维 hiring trends over the years

运维 changes in hiring volume over the years

深圳 运维 related job postings

Based on related job postings of 深圳市工匠行科技有限公司 in the past year
  • 大模型运维工程师

    深圳-南山区 | 1-3年 | 本科以上 | 2026-09-17
    16000-20000
    岗位职责
    1、本地大模型部署与优化:使用 vLLM/SGLang 部署Qwen3.6 等大模型,配置张量并行、流水线并行,实现高吞吐推理服务。
    2、模型量化与压缩:采用 AWQ/GPTQ/FP8/INT4 等量化技术,将大模型压缩适配到本地硬件(如单卡 24GB 显存运行 32B 模型),平衡性能与资源占用,满足边缘设备及低资源场景需求。
    3、推理服务化建设:搭建兼容 OpenAI 协议的 API 服务,为开发团队提供标准化调用接口,实现 IDE 插件、AI 工具链及公司内部业务系统的集成对接。
    4、性能调优与稳定性保障:优化推理延迟(TTFT/TPOT),建立显存监控与告警机制,确保多用户并发使用时的服务稳定性与高可用性。
    任职要求
    1、计算机相关专业本科及以上学历,1-3年 Linux 系统运维或 DevOps 经验。
    2、精通 Docker/K8s 容器化部署,熟悉 CUDA 环境配置与 GPU 驱动管理(RTX 4090/A100/H100 等)。
    3、掌握 vLLM、llama.cpp、Ollama 等主流推理框架,具备 70B + 参数大模型本地部署的实战经验。
    4、熟悉 AWQ/GPTQ/GGUF 等模型量化技术,能根据硬件条件制定*优量化与部署方案。
    5、了解 Triton Inference Server 或 NVIDIA TensorRT-LLM 等推理加速工具,有实际优化经验优先。
    6、了解 GPU 互联技术(NVLink/PCIe),能排查显存溢出(OOM)与通信瓶颈等常见问题。
    7、熟悉内网环境下的模型服务防护,能配置 API 访问控制(JWT 认证 + 限流),保障服务安全。
    加分项
    1、有边缘端 AI 模型部署、轻量模型蒸馏 / 压缩经验优先。
    2、有具身智能、嵌入式设备 AI 适配经验(如轮椅、智能硬件场景)优先。
    3、熟悉多模型混合部署、多租户隔离与资源调度方案。
    4、有大模型推理性能压测、高并发调优及线上故障处理经验。
    More

深圳 运维 salary

More
How much does 深圳 运维 pay? 10-15K is the most common