cs.LG 편의 논문 | Gist.Science

An accurate flatness measure to estimate the generalization performance of CNN models

이 논문은 완전 연결 네트워크에 국한되거나 근사적인 기존 방법의 한계를 극복하기 위해, 합성곱 신경망 (CNN) 의 기하학적 구조를 정확히 반영하는 폐쇄형 평탄도 측정치를 제안하고 이를 통해 CNN 모델의 일반화 성능을 정밀하게 평가하고 아키텍처 설계에 활용할 수 있음을 입증합니다.

Rahman Taleghani, Maryam Mohammadi, Francesco Marchetti2026-03-11🤖 cs.LG

When to Retrain after Drift: A Data-Only Test of Post-Drift Data Size Sufficiency

이 논문은 개념 변화 (concept drift) 발생 후 재학습에 필요한 충분한 데이터 크기를 결정하기 위해, 동적 시스템의 상태 의존성을 활용한 단일 패스 가중 국소 회귀 기반의 데이터 전용 테스트 'CALIPER'를 제안하고 다양한 도메인에서 그 유효성을 입증합니다.

Ren Fujiwara, Yasuko Matsubara, Yasushi Sakurai2026-03-11🤖 cs.LG

Two Teachers Better Than One: Hardware-Physics Co-Guided Distributed Scientific Machine Learning

이 논문은 중앙 집중식 처리의 한계를 극복하기 위해 하드웨어와 물리 법칙을 공동으로 안내하는 분산 과학 머신러닝 프레임워크 'EPIC'을 제안하여, 경량 인코딩과 물리 인식 디코딩을 통해 통신 지연과 에너지 소모를 획기적으로 줄이면서도 물리적 정밀도를 유지하거나 향상시킨다는 점을 보여줍니다.

Yuchen Yuan, Junhuan Yang, Hao Wan, Yipei Liu, Hanhan Wu, Youzuo Lin, Lei Yang2026-03-11🤖 cs.LG

SCALAR: Learning and Composing Skills through LLM Guided Symbolic Planning and Deep RL Grounding

이 논문은 LLM 기반 계획과 강화학습을 양방향으로 결합하여 실행 피드백을 통해 기술 명세를 반복적으로 정제하는 'SCALAR' 프레임워크를 제안하며, Craftax 환경에서 기존 최선 방법론 대비 1.9 배 향상된 성능을 입증했습니다.

Renos Zabounidis, Yue Wu, Simon Stepputtis, Woojun Kim, Yuanzhi Li, Tom Mitchell, Katia Sycara2026-03-11🤖 cs.LG

FlexServe: A Fast and Secure LLM Serving System for Mobile Devices with Flexible Resource Isolation

이 논문은 ARM TrustZone 의 비효율적인 리소스 격리 문제를 해결하기 위해 유연한 메모리 및 NPU 격리 메커니즘을 도입하여 모바일 기기에서 LLM 추론 속도와 보안을 동시에 극대화하는 'FlexServe' 시스템을 제안합니다.

Yinpeng Wu, Yitong Chen, Lixiang Wang, Jinyu Gu, Zhichao Hua, Yubin Xia2026-03-11🤖 cs.LG

From Days to Minutes: An Autonomous AI Agent Achieves Reliable Clinical Triage in Remote Patient Monitoring

본 논문은 원격 환자 모니터링 데이터를 실시간으로 분석하여 개별 임상진료자보다 높은 민감도로 응급 상황을 식별하고, 확장 가능한 비용 효율적인 임상 분류를 가능하게 하는 자율 AI 에이전트 'Sentinel'의 개발과 유효성을 입증했습니다.

Sim2Act: Robust Simulation-to-Decision Learning via Adversarial Calibration and Group-Relative Perturbation

이 논문은 시뮬레이션 오차를 의사결정 영향도에 따라 재가중하는 적대적 보정 메커니즘과 시뮬레이션 불확실성 하에서 정책 학습을 안정화하는 그룹 상대적 교란 전략을 통해, 공급망 등 임무 중대 분야에서 견고한 시뮬레이션-의사결정 학습 프레임워크인 Sim2Act 를 제안합니다.

Hongyu Cao, Jinghan Zhang, Kunpeng Liu, Dongjie Wang, Feng Xia, Haifeng Chen, Xiaohua Hu, Yanjie Fu2026-03-11🤖 cs.AI

Quality over Quantity: Demonstration Curation via Influence Functions for Data-Centric Robot Learning

이 논문은 로봇 학습의 성능을 높이기 위해 검증 데이터의 손실 감소에 기여하는 정도를 기반으로 각 시연 데이터의 품질을 정량화하고, 영향 함수 (influence functions) 를 활용해 고품질 데이터를 체계적으로 선별하는 'Quality over Quantity (QoQ)' 방법을 제안합니다.

Haeone Lee, Taywon Min, Junsu Kim, Sinjae Kang, Fangchen Liu, Lerrel Pinto, Kimin Lee2026-03-11🤖 cs.LG

Adaptive Active Learning for Online Reliability Prediction of Satellite Electronics

본 논문은 제한된 데이터와 다양한 운영 조건 하에서 위성 전자기기의 온궤도 신뢰성 예측 정확도를 향상시키기 위해, Wiener 과정 기반의 고장 모델과 공간적 상관관계를 통합한 적응형 능동 학습 프레임워크를 제안합니다.

Shixiang Li, Yubin Tian, Dianpeng Wang, Piao Chen, Mengying Ren2026-03-11🤖 cs.LG

Dynamic Multi-period Experts for Online Time Series Forecasting

이 논문은 개념 변화 (Concept Drift) 를 재발생과 신규 발생으로 재정의하고, 각각에 맞춰 역사적 패턴을 활용하거나 안정적인 일반 전문가로 전환하는 'DynaME'라는 새로운 하이브리드 프레임워크를 제안하여 온라인 시계열 예측 성능을 크게 향상시킵니다.

Seungha Hong, Sukang Chae, Suyeon Kim, Sanghwan Jang, Hwanjo Yu2026-03-11🤖 cs.LG

Learning Adaptive LLM Decoding

이 논문은 고정된 샘플링 하이퍼파라미터 대신 강화학습을 통해 추론 시 계산 자원에 따라 동적으로 샘플링 전략을 선택하는 경량 디코딩 어댑터를 제안하여, 수학 및 코딩 벤치마크에서 고정된 예산 대비 정확도를 크게 향상시킨다는 점을 설명합니다.

Chloe H. Su, Zhe Ye, Samuel Tenka, Aidan Yang, Soonho Kong, Udaya Ghai2026-03-11🤖 cs.LG

Verifying Good Regulator Conditions for Hypergraph Observers: Natural Gradient Learning from Causal Invariance via Established Theorems

이 논문은 Wolfram 의 초그래프 물리학과 Vanchurin 의 신경망 우주론을 기반으로, 인과 불변 초그래프 기반의 지속적 관찰자가 Conant-Ashby 좋은 조절자 정리를 만족하고 자연 기울기 하강법이 유일한 학습 규칙임을 증명하며, 이를 통해 다양한 수렴 모델에 따라 관찰자가 피셔 계량 텐서의 고유 방향을 따라 서로 다른 Vanchurin 체제에 동시에 존재할 수 있음을 규명합니다.

Max Zhuravlev2026-03-11🤖 cs.LG

← 이전 다음 →

cs.LG

An accurate flatness measure to estimate the generalization performance of CNN models

When to Retrain after Drift: A Data-Only Test of Post-Drift Data Size Sufficiency

Two Teachers Better Than One: Hardware-Physics Co-Guided Distributed Scientific Machine Learning

SCALAR: Learning and Composing Skills through LLM Guided Symbolic Planning and Deep RL Grounding

FlexServe: A Fast and Secure LLM Serving System for Mobile Devices with Flexible Resource Isolation

From Days to Minutes: An Autonomous AI Agent Achieves Reliable Clinical Triage in Remote Patient Monitoring

Sim2Act: Robust Simulation-to-Decision Learning via Adversarial Calibration and Group-Relative Perturbation

Quality over Quantity: Demonstration Curation via Influence Functions for Data-Centric Robot Learning

Adaptive Active Learning for Online Reliability Prediction of Satellite Electronics

Dynamic Multi-period Experts for Online Time Series Forecasting

Learning Adaptive LLM Decoding

Verifying Good Regulator Conditions for Hypergraph Observers: Natural Gradient Learning from Causal Invariance via Established Theorems

Exclusive Self Attention

PPO-Based Hybrid Optimization for RIS-Assisted Semantic Vehicular Edge Computing

Not All News Is Equal: Topic- and Event-Conditional Sentiment from Finetuned LLMs for Aluminum Price Forecasting

Latent World Models for Automated Driving: A Unified Taxonomy, Evaluation Framework, and Open Challenges

Overcoming Valid Action Suppression in Unmasked Policy Gradient Algorithms

Probabilistic Hysteresis Factor Prediction for Electric Vehicle Batteries with Graphite Anodes Containing Silicon

Decoupling Reasoning and Confidence: Resurrecting Calibration in Reinforcement Learning from Verifiable Rewards

Causally Sufficient and Necessary Feature Expansion for Class-Incremental Learning