гросссофт

Архитектор решений(LLM)

Вилка: от 1500 до 2000 в час (рубли) без НДС - почасовая оплата

гросссофт На сервисе с: 05.10.26 15:04

240k–320k ₽РоссияУдалёнка

Архитектор решений

Локация: РФ

Вилка: от 1500 до 2000 в час (рубли) без НДС - почасовая оплата

Работа по договору с ИП (статус ИП у вас)

Архитектурный и инфраструктурный стек:

  • Kubernetes, Docker, Helm, Nginx/Ingress, Kafka/RabbitMQ, Redis, PostgreSQL, Object Storage, API Gateway, Vault, Keycloak/ADFS, Prometheus, Grafana, OpenSearch.

ML / MLOps стек:

  • Python, FastAPI, vLLM, Triton, PyTorch, MLflow, ClearML, Airflow, GitLab CI/CD.

Интеграционный стек:

  • REST, WebSocket, OpenAPI, AsyncAPI, ETL, корпоративные IAM/SSO, LDAP/AD.

Требования:

  • Архитектура highload-сервисов: Опыт проектирования высоконагруженных распределённых систем: горизонтальное масштабирование, отказоустойчивость, балансировка нагрузки, кэширование, очереди, rate limiting, graceful degradation.
  • Архитектура ML/LLM-сервисов: Опыт проектирования архитектуры LLM-сервисов для промышленной эксплуатации: inference-serving, batch/online processing, fvector storage, orchestration пайплайнов, GPU workload management.
  • Интеграция в корпоративный ИТ-ландшафт: Опыт встраивания ML-сервисов в корпоративную архитектуру: интеграция с API Gateway, IAM/SSO, корпоративными шинами, системами мониторинга и журналирования.
  • Микросервисная и событийная архитектура: Понимание микросервисного подхода, SOA, event-driven architecture, async-взаимодействия, брокеров сообщений и потоковой обработки данных.
  • Контейнеризация и оркестрация: Уверенное знание Docker, Kubernetes, Helm, ingress/service mesh, принципов деплоя и эксплуатации контейнеризированных решений.
  • Облачная и on-prem архитектура: Опыт проектирования решений как в on-prem, так и в private/public cloud, понимание hybrid architecture, multi-zone deployment, disaster recovery, backup/restore, SLA/SLO/SLI.
  • Интеграционные технологии и API: Опыт проектирования REST/WebSocket интеграций, контрактов взаимодействия, схем авторизации и маршрутизации.
  • Безопасность и корпоративные стандарты: Знание принципов ИБ: TLS/mTLS, JWT, OAuth2, OIDC, SAML/ADFS, RBAC/ABAC, secrets management, требования к защите ПДн и чувствительных данных.
  • Данные и хранилища: Понимание архитектуры хранения и обработки данных: PostgreSQL/Redis/Object Storage.
  • Observability и эксплуатация: Опыт проектирования мониторинга и сопровождения сервисов: метрики, логи, трассировки, алерты, capacity planning, performance profiling, incident/problem management. MLOps / DevOps практики: Понимание CI/CD для ML-сервисов, model registry, experiment tracking, reproducibility.
  • Технический стек и языки: Понимание Python, SQL, Bash, Git. Умение читать код и проектировать архитектуру с учётом ограничений фреймворков и runtime.
  • Архитектурная документация: Опыт подготовки архитектурных артефактов: C4/UML, схемы интеграции, NFR, ADR, security/data flow diagrams.
  • Опыт и уровень: Практический опыт работы системным/solution/enterprise architect от 3х лет, общий опыт в ИТ от 5 лет. Опыт с LLM-продуктами или платформами данных.

Дополнительно:

  • Опыт проектирования GPU-инфраструктуры и inference-контуров для LLM/embedding/reranking сервисов.
  • Понимание особенностей vLLM, Triton Inference Server, ONNX Runtime, MLflow, ClearML.
  • Опыт выбора архитектуры под разные профили нагрузки: low latency / high throughput / batch / streaming.
  • Навык оценки стоимости владения архитектурой: CAPEX/OPEX, стоимость GPU/CPU/Storage/Network.
  • Опыт проектирования мультитенантных платформ и контуров изоляции.
  • Знание практик FinOps / Capacity planning будет плюсом.
  • Опыт прохождения архитектурных комитетов, согласования решений с ИБ, инфраструктурой и корпоративными архитекторами.
  • Навык формализации нефункциональных требований: производительность, доступность, безопасность, сопровождаемость, масштабируемость.
  • Понимание специфики корпоративной интеграции: legacy-системы, ограниченные контуры.

Похожие вакансии Системный архитектор

Другие площадки
raft digital solutions

AI архитектор

raft digital solutionsНа сервисе с: 17.07.26 21:32↑ Вакансия с автоподнятием
400k–588k ₽Не указана странаЛокация не указанаУдалёнка

Привет! Мы - команда Raft Digital Solutions, занимаемся разработкой решений на базе AI, внесли свой вклад во фреймворк Langchain, создали собственный инновационный продукт для анализа голосовой связи с помощью GPT, а также провели обширные исследования и разработки в области безопасности LLM. Работаем как на рынке РФ, так и на международном.

Сейчас мы в поиске AI архитектора в наше зарубежное направление.

! Важна локация кандидатов вне РФ и вне РБ, а также наличие разговорного английского.

Чем предстоит заниматься:

  • Пресейлы с потенциальными клиентами
  • Подготовка эстимейтов и пропоузалов на скоуп работ
  • Прохождение клиентских интервью для старта проектов
  • Ведение технической части в проектах: помощь разработчикам в построении изначального решения
  • Контроль выполнения работ инженерами и попадание в сроки в соответствии с эстимейтами

Мы ожидаем от кандидатов:

  • Опыт на позиции AI архитектора или аналогичной роли
  • Широкий опыт работы с клиентами (преимущественно со стартапами)
  • Опыт подготовки оценок для проектов, организация работы и контроль попадания в сроки
  • Важен бекграунд в backend разработке, знание Python на уровне senior, опыт с современными библиотеками и фреймворками (особенно FastAPI), опыт проведения код-ревью
  • Коммерческий опыт с GenAI, интеграции AI решений в приложения (LLM, RAG, ИИ-агенты и др.)
  • Знание популярных инструментов для интеграции AI/LLM в приложение (например, Hugging Face Transformers, LangChain, OpenAI Python SDK, PydanticAI, LlamaIndex, Microsoft AutoGen, Ollama и др.)
  • Знание и опыт с CI/CD, DevOps, MLOps
  • Если есть экспертиза в конкретном домене в AI (Healthcare, Video, etc.), будет огромным плюсом
  • Разговорный английский язык на уровне не ниже B2 (Upper-Intermediate)
  • ! Ваша локация должна быть вне РФ, вне РБ

Мы предлагаем:

  • Удалённый формат работы
  • Гибкий рабочий график, фулл-тайм
  • Задачи на стыке бизнеса и современных AI-технологий
  • Возможность получить широкую экспертизу в разнообразных доменах бизнеса
  • Влияние на развитие продуктов и на технические решения
  • Дружную команду экспертов, обмен знаниями и тех.рост
  • Заключаем договор формата b2b (с ИП, валютный расчётный счёт)
Эта вакансия также есть на:Hirify
Сайты компаний
N

Partner Solutions Architect

nebiusНа сервисе с: 06.10.26 00:36↑ Вакансия с автоподнятием
Зарплата не указанаRemote - United StatesУдалёнка

About Nebius:

Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure.

Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI.

Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D.

Customer experience:

Customer experience at Nebius AI Cloud involves tackling customers’ challenges and directly impacting their success by solving real-world AI and ML problems at massive GPU cloud scale. You’ll not only resolve issues, but play a key role in shaping clients’ business success by optimizing their AI solutions. Working with advanced GPUs such as H200, B200 and GB200, as well as modern ML frameworks, you’ll influence the development of the Nebius AI Cloud and gain experience at the intersection of infrastructure and AI. With minimal bureaucracy, you’ll have the freedom to innovate, take ownership and drive change. Opportunities for growth are abundant in this vibrant and supportive professional community.

The role

We are looking for a Partner Solutions Architect to serve as the technical interface between Nebius and our strategic technology partners, ranging from data platforms and MLOps vendors to AI frameworks and ISVs that build on GPU infrastructure.

This is a hands-on engineering and integration role. You will design and develop integrated solutions, build reference implementations, enable partner engineering teams, and ensure joint customers succeed with combined offerings. You will influence Nebius’ product roadmap and drive deep technical collaboration across partner ecosystems.

You’re welcome to work remotely from the United States or Canada.

Your responsibilities will include: 

  • Own the technical relationship with strategic partners as the primary interface to Nebius engineering and product
  • Design and architect high-impact integration solutions, delivering clear joint customer value
  • Build and maintain reference implementations and production-grade integrations
  • Define enterprise-ready integration patterns (networking, security, compliance, observability)
  • Translate partner and customer feedback into actionable product and roadmap requirements
  • Enable partner teams through technical documentation, training, and demo environments
  • Provide technical leadership in joint customer deployments, architecture reviews, and complex troubleshooting

We expect you to have: 

  • 7+ years in solutions architecture, partner engineering, or technical pre-sales at cloud or data platform companies
  • Bachelor’s degree or foreign equivalent in a related field, or an equivalent combination of education and relevant experience.
  • Deep cloud infrastructure expertise, with specialist experience in large-scale Kubernetes deployments and containerized environments
  • Strong hands-on experience with Infrastructure as Code, especially Terraform
  • Hands-on experience deploying production AI workloads, particularly inference, and architecting or implementing agentic solutions
  • Proven background in integration architecture, including API design, data pipelines, security, and observability, with experience in Python
  • Excellent communication skills with the ability to engage technical and non-technical

It will be an added bonus if you have: 

  • Experience with HPC environments and large-scale GPU clusters
  • Hands-on experience across the ML lifecycle, including training, fine-tuning, and RAG
  • Practical experience building AI agents, particularly with LangChain
  • Experience with partner technical programs, major cloud or data platforms, or enterprise compliance standards
  • Deep understanding of advanced Kubernetes infrastructure concepts beyond the abstraction of managed Kubernetes services.

Key Employee Benefits:

  • Health Insurance: 100% company-paid medical, dental, and vision coverage for employees and families.
  • 401(k) Plan: Up to 4% company match with immediate vesting.
  • Parental Leave: 20 weeks paid for primary caregivers, 12 weeks for secondary caregivers.
  • Remote Work Reimbursement: Up to $85/month for mobile and internet.
  • Disability & Life Insurance: Company-paid short-term, long-term, and life insurance coverage.

Join Nebius Today!

Pay Transparency

We offer competitive compensation and benefits packages. Actual compensation will be determined based on job-related factors, including experience, skills, qualifications, the level at which the candidate is hired, and geographic location, consistent with applicable law.

On Target Earnings Range
$250,000—$320,000 USD

Benefits & Perks:

  • Competitive compensation
  • Career growth and learning opportunities
  • Flexibility and ownership
  • Collaborative and innovative culture
  • Opportunity to work on impactful AI projects
  • International environment and talented teams

What's it like to work at Nebius:

Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI 

Equal Opportunity Statement:

Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law.

Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire. 

If you need accommodations during the application process, please let us know.

Другие площадки
B

Solutions Architect, EMEA

basetenНа сервисе с: 09.10.26 18:17
Зарплата не указанаEMEAУдалёнка

ABOUT BASETEN

Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F, led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products.

THE ROLE
As a Solutions Architect at Baseten you will partner closely with Sales and customers to translate business needs into technical solutions, run technical discovery, and guide repeatable deployments and proofs of value for customers. This role is a great fit for entrepreneurial, customer-facing technical professionals who want a front-row view into how modern companies adopt AI at scale, and who enjoy working across technical discovery, solution design, demos, deployment scoping, and hands-on customer implementations, in close partnership with Sales and Engineering.


RESPONSIBILITIES

  • Partner with Sales on customer discovery calls (most often second calls, occasionally first calls for large accounts).

  • Lead demos and technical scoping to align on success criteria, architecture, and deployment approach.

  • Own benchmarking and repeatable deployments, including:

  • Handling standard deployment patterns and configurations across many modalities – LLMs, embeddings, image and video generation, Voice AI, etc.

  • Advising on tradeoffs like H100s vs B200s and latency-optimized vs throughput-optimized setups.

  • Driving consistent “playbook” style deployments for common models and use cases.

  • Become a power user of different runtimes such as vLLM, SGLang, and TRT-LLM and all the common configurations and tradeoffs between them

  • Drive POC and project execution, including:

  • Scoping POCs and keeping stakeholders aligned on timeline, deliverables, and next steps.

  • Acting as the “ringleader” or project manager for POCs.

  • Pulling in Forward Deployed Engineering (FDE) support when deeper or more complex technical work is needed.

REQUIREMENTS

  • AI/ML background and the ability to credibly discuss AI/ML topics with technical stakeholders.

  • Strong customer-facing communication skills, including the ability to run structured discovery and clarify ambiguous requirements.

  • Technical depth to scope solutions, without needing to write production code.

  • Ability to script and prototype as needed, including comfort “vibe coding” to move quickly in technical workflows.

NICE TO HAVE

  • Experience running or supporting benchmarks for ML inference deployments.

  • Familiarity with infrastructure tradeoffs relevant to inference performance and cost (for example GPU selection and latency versus throughput tuning).

  • Experience serving as a cross-functional technical lead for customer POCs, including coordination across Sales and Engineering.

BENEFITS

  • Competitive compensation, including meaningful equity

  • (U.S. only) 100% coverage of medical, dental, and vision insurance for employee and dependents

  • Flexible PTO policy including company wide Winter Break (our offices are closed from Christmas Eve to New Year's Day!)

  • Paid parental leave

  • Fertility and family-building stipend through Carrot

  • (U.S. only) Company-facilitated 401(k)

  • Exposure to a variety of ML startups, offering unparalleled learning and networking opportunities.

Apply now to embark on a rewarding journey in shaping the future of AI! If you are a motivated individual with a passion for machine learning and a desire to be part of a collaborative and forward-thinking team, we would love to hear from you.

At Baseten, we are committed to fostering a diverse and inclusive workplace. We provide equal employment opportunities to all employees and applicants without regard to race, color, religion, gender, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, or veteran status.

We are an Equal Opportunity Employer and will consider qualified applicants with criminal histories in a manner consistent with applicable law (by example, the requirements of the San Francisco Fair Chance Ordinance, where applicable).

HireSeeker собирает вакансии со всех площадок и присылает только релевантные. Бесплатно.