Accepted at NeurIPS 2026

SLM-Agents

1st Workshop on SLMs for Agentic Systems

Paris, France · December 12–13, 2026

Exact workshop day to be announced

About the workshop

Scope and objectives

This workshop is dedicated to small language models (SLMs) as the foundation of agentic AI systems. Although large language models (LLMs) have demonstrated remarkable capabilities, their dependence on cloud infrastructure creates fundamental barriers to deployment in agentic pipelines—latency, privacy, connectivity, and substantial computational cost. SLMs offer a compelling alternative: recent studies argue that SLMs, not LLMs, might be a right option for the repetitive, narrowly scoped sub-tasks that dominate real agentic workloads. SLMs make it possible for autonomous AI agents to plan, reason, and act directly on resource-constrained devices such as smartphones, IoT systems, robotics platforms, and embedded systems. The workshop sits at the intersection of three rapidly evolving fields: (1) efficient language model architectures and compression techniques, (2) agentic AI systems capable of autonomous reasoning and tool use, and (3) edge computing and on-device deployment.

Open problems

  • Compression and distillation: quantization, pruning, knowledge distillation, and architectural innovations for parameter-efficient LMs.
  • Hardware co-design: on-device inference optimization, NPU/accelerator-aware design, memory-bandwidth-bound serving, and energy-budgeted decoding.
  • Training for cooperation: fine-tuning SLMs for tool use, planning, multi-step reasoning, and small–large model handoff in heterogeneous agent stacks.
  • Evaluation and benchmarks: task-success-per-watt, latency- and memory-aware leaderboards, and reproducible on-device evaluation harnesses.
  • Applications and safety: privacy-preserving local processing, federated learning, and deployment case studies across mobile assistants, robotics, healthcare, automotive, and financial services, with associated safety, robustness, and provenance considerations.

Research contributions

Call for Papers

We invite original work in progress on small language models for agentic systems.

Submission format

  • Extended abstracts: 4 pages plus references
  • Optional full papers: 8 pages plus references
  • NeurIPS workshop template
  • Double-blind review through OpenReview
  • Three reviewers per submission

Scope and publication

  • Original work in progress
  • Not under review at or accepted to the NeurIPS 2026 main program
  • Not previously published at a major ML or AI venue
  • Accepted papers hosted on OpenReview with author opt-in
  • Non-archival and not included in the NeurIPS proceedings

Topics of interest

  • SLM architectures, training and inference
  • Agentic reasoning, planning and tool use
  • Hardware-aware and on-device deployment
  • Evaluation, benchmarks and efficiency
  • Privacy, safety, robustness and applications

Submission portal coming soon

Submission timeline

Important Dates

Call for Papers released

July 18, 2026

Submission portal opens

July 25, 2026

Paper submission deadline

August 22, 2026 (AoE)

Reviewing period

August 30–September 19, 2026

Acceptance notification

September 22, 2026

Camera-ready deadline

October 13, 2026

Workshop

December 12 or 13, 2026

Featured talks

Invited Speakers

Portrait of Emmanuel Abbe

Emmanuel Abbe

EPFL and Apple

Portrait of Ali Ghodsi

Ali Ghodsi

University of Waterloo

Portrait of Diana Marculescu

Diana Marculescu

The University of Texas at Austin

Portrait of Amir Gholami

Amir Gholami

University of California, Berkeley

Community discussion

Invited Panelists

Invited panelists will be announced once confirmed.

One day · In person · Paris

Workshop Schedule

Workshop schedule: TBD.

Companion activity

Edge Agent Efficiency Challenge

The challenge focuses on running a multi-step agent task on consumer-class hardware, such as a laptop GPU, mobile NPU, or single-board computer, under fixed memory, latency, and energy budgets.

Entries will be evaluated using a cost-adjusted task-success metric and will release reusable model checkpoints and inference recipes. Challenge winners will present their systems during the workshop’s live-demo session.

Workshop leadership

Organizers

Portrait of Habib Hajimolahoseini

Organizer

Habib Hajimolahoseini

AMD

Portrait of Mehdi Rezagholizadeh

Organizer

Mehdi Rezagholizadeh

AMD

Portrait of Vahid Partovi Nia

Organizer

Vahid Partovi Nia

École Polytechnique de Montréal and Huawei Canada Noah’s Ark Lab

Portrait of MirHamed Jafarzadeh Asl

Organizer

MirHamed Jafarzadeh Asl

Huawei Noah’s Ark Lab

Portrait of Shahrzad Kianidehkordi

Organizer

Shahrzad Kianidehkordi

RBC Borealis

Portrait of Mouloud Belbahri

Organizer

Mouloud Belbahri

TD Insurance

Portrait of Pavlo Molchanov

Organizer

Pavlo Molchanov

NVIDIA Research

Peer review

Scientific Committee

The scientific committee will be announced once confirmed.

Review outcomes

Accepted Papers

Accepted papers will be listed after the review process is complete.

View accepted papers

Participation

Policies, Accessibility, and Inclusion

Non-archival workshop

Accepted workshop submissions are non-archival and may be submitted elsewhere after the workshop, subject to the policies of the other venue.

Participation and inclusion

The program includes mentorship for junior researchers, a dedicated lightning-talk track, outreach through affinity communities, and an inclusive workshop environment for participants across backgrounds and career stages.

Stay informed

Contact and updates

For questions, contact the organizers at neurips.slmagents.2026@gmail.com. Last updated July 23, 2026.