ARYAN KARGWAL

[PhD Researcher] AI & Robotics
@ Polytechnique Montréal

Aryan Kargwal

[ THE LAB BENCH ]

Systems Engineering & Research Prototypes

Active Development

BenchYard: Edge Device Benchmarking & LLM Inferencing

Edge devices benchmarking and inferencing for LLMs. Optimizing large language model deployment and performance evaluation across edge computing environments.

LLMEdge ComputingBenchmarking
Research Prototype

Multi-Robot SLAM: Formal Language Verification

Formal language verification approach for multi-agent SLAM systems. Uses temporal logic specifications to ensure correctness guarantees in decentralized mapping and localization tasks.

RoboticsSLAMFormal Verification
Archived

Caiman Ice: Semantic Segmentation for Remote Sensing

SLAM-based semantic segmentation for ice and snow classification in remote sensing imagery. Uses SLAM-derived super-segments merged via visual cues to distinguish between water, snow, and ice in ambiguous conditions.

Semantic SegmentationRemote SensingSLAM
#01

Agent evaluation metrics: how to measure whether an agent works

Arize AI May 2026

Learn how to measure and evaluate AI agent performance effectively with key metrics and methodologies.

Agent Ops
p}]]W*>] ;jl(UQ_:jovMO'C\8zo(0Q_
IJ\a^nk'_mYvnpjjqh!oQox.8[Y|qipa
+{I?mW~czO<c@#_@J]xB>_vI"n`"M?@~
X(-d YYjU[\jwof\>@BcY`pa|(vwoo''
^_0wn)|jvlxZ,x ?^W1UQ~M|qZ-[aatJ
M*!U~l>U`)]_)^(;&^/tZl_p}:qW1/1~
+;ah@a'||lB{wU]]{W M^w@B`-O^ L/^
[fa|j}%m<qpkJ$]QM|CC/[&%o`Mr0)\0

BETA DISTRIBUTION

#02

Best AI Observability Tools for Autonomous Agents in 2026

Arize AI Feb 2026

A comparative look at leading observability tools for autonomous AI agents, with evaluation criteria for production deployment in 2026.

Tools & Platforms
?i!i">>i+_+{+!-[+_+':+"><!,`.><!
i;l>l,>li!l>I]!I:>~>iI<`:>>+<:i`
.->-,;^:!_}[f|)\xX|(~1)<}<,><"~!
?,+-l??+}\njzOCJLqpUcf|[1~[_<i I
!_><{!}/Xruj#*aB$kZZ0zUtv/[?}i><
l?>]-?<]}jr|QLzqY0Czr|Y1?[?!>!?<
^_I~~+1:]+1]rjr|ff~\<[~|+-<>~i+?
> i>>{,il[~<!+-+?<1?<l+]~:~l><,l

GAUSSIAN DISTRIBUTION

#03

Understanding Latency in AI Model Deployment

Pipeshift Feb 2026

A practical breakdown of where latency comes from in deployed AI systems and how teams can reduce bottlenecks across the inference path.

Engineering
 ^ ' '.`' '''`''` .'`'. '`'`  ''
-+ii;,;^`"^`.^^ .^' `^  .`'`` '^
v/[]>!;l;"`'^"^'`^.`^`.'.` `'.'`
bCn\]-<>:;"""`^^`..``.. `'^.^'^'
$kUx\[?_li;"^,`'```'^^' ^..  . .
wCr\]]+<Il:;`^'''^'.`' .'``  ^  
n|{+~i;I";^`"`^`^..'`'  `   ` ` 
_+>I:I;^""^.`` .. .^.``'^`^ .^  

EXPONENTIAL DECAY

#04

Rented AI and Hidden Costs

Pipeshift Feb 2026

An analysis of the hidden operational and business costs teams face when relying on rented AI stacks without strong infrastructure ownership.

Strategy
'. '`' ''^'.'.`^  ^^ '`'.   ''.^
-~~!I;,","'^^` ^''^'` .'  `' '.'
c/{]<~!I;,^``''.^ '^`'`..'^  ` .
pUn/}++!!lI"````^.^^^..''`.`^ '^
$aQx\}_~>II,":"^''''. '^ ^''.''.
dYr/[-~i!:,":'^`'""^^ `.''. '^.'
cf{_~ilI,;,"'`^.^` ^'` ''`. `..'
]<!!l"^::'^^`'`'.^' ^`.`...'`'`.

EXPONENTIAL DECAY

#05

Why AI Agents Break: A Field Analysis of Production Failures

Arize AI Jan 2026

A comprehensive field analysis of production failures in AI agents. Explores the root causes and failure patterns that occur when AI agents break in real-world deployments.

Analysis
`..`.  ''.^^^::;II;I;^"'`...''`'
'`.. ''^`.",I_~]1}}]~il,"`'' ``.
. .''``.""l;+/\cUYcct]-I`".'`.' 
'' '.'.^,,i<1XXb*adpX()>,^^^.` '
  .. '``;I~<\OLM$%##0/|~;I`^`''.
 '.'. '`",>i|UYd*odbY((i;;^'.'. 
  '...'.`"ll+||vYYuu\?-!^``'`'`.
 . `  .^^'":!_~]{[?}+ll"`^.` `..

BINOMIAL DISTRIBUTION

#06

Observability-Driven Sandboxing

Arize AI Jan 2026

Securing AI agents through observability-driven sandboxing: techniques for monitoring and controlling agent behavior.

Security
`'.`. ``^''.' ..`'``'``' '. `'..
?_<i:",`,`"'. .  . '. .''`'. `. 
v([]~>;:;:^`^'"." '` `` `. ''. ^
mXx(]?<!!I,"^,.^`^'^` .`.^  `.' 
$kJr\1_<<!lI^"^^.'..`^ ^`` `.^'.
mUx1}_<ll:I:"^"".`.`^`. .`^..`^ 
v({]<l!I;;^''`'''`^` ` `^'' . .`
]<iI,",`^'^..^.^.`.^. `'^. ..`.`

EXPONENTIAL DECAY

[ Currently Reading ]

View on Goodreads →