[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"doc-detail-81602-en":3,"doc-seo-81602-105":30,"detail-sidebar-cat-0-en-105":91},{"code":4,"msg":5,"data":6},0,"success",{"doc_id":7,"user_id":8,"nickname":9,"user_avatar":10,"doc_module":4,"category_id":11,"category_name":12,"doc_title":13,"doc_description":14,"doc_content":15,"file_id":16,"file_url":17,"file_type":18,"file_size":19,"view_count":20,"is_deleted":4,"is_public":21,"is_downloadable":21,"audit_status":21,"page_count":22,"language":23,"language_code":24,"site_id":25,"html_lang":24,"table_of_contents":26,"faqs":27,"seo_title":13,"seo_description":14,"update_tm":28,"read_time":29},81602,34359740700684,"Finn","https://ap-avatar.wpscdn.com/avatar/1f400023980c374ae676?_k=1777273430885731487",8,"Research & Report","Fine-grained Soundscape Control for Augmented Hearing","Hearables are widely used, yet their sound controls remain limited to global noise suppression or single-target focus. Real acoustic scenes contain many simultaneous sources that users may wish to adjust independently. Aurchestra is introduced as a first fine-grained, real-time soundscape control system for resource-constrained hearables, combining a dynamic interface that exposes active sound classes with an on-device multi-output extraction network that streams each selected class. The system supports up to five overlapping targets, enabling per-class volume mixing on 6 ms audio chunks and delivering robust enhancement and interference suppression in unseen indoor and outdoor environments.","arXiv :2603 .00395v 3 [ cs . SD] 10 Jul 2026  \nFine-grained Soundscape Control for Augmented Hearing  \nSeunghyun Oh, 1 Malek Itani, 1,2 Aseem Gauri, 1 Shyamnath Gollakota 1,2  \n1Paul G. Allen School of Computer Science and Engineering, University of Washington  \n2Hearvana AI  \nFigure 1: Aurchestra transforms the auditory world into a programmable studio. Unlike traditional hearables that offer binary noise cancellation (all-or-nothing), Aurchestra enables fine-grained soundscape control. (A) In a complex acoustic scene,(B) the system automatically detects active sound classes (e.g., speech, traffic, birds) and populates a dynamic interface. The user can then “mix” their reality in real-time, (C) independently suppressing interference (traffic) while increasing the volume of some targets (speech) and maintaining others (nature), effectively acting as the audio engineer of their own life.  \nAbstract  \nHearables are becoming ubiquitous, yet their sound controls remain blunt: users can either enable global noise suppression or focus on a single target sound. Real-world acoustic scenes, however, contain many simultaneous sources that users may want to adjust independently. We introduce Aurchestra, the first system to provide fine-grained, real-time soundscape control on resourceconstrained hearables. Our system has two key components: (1) a dynamic interface that surfaces only active sound classes and (2) a real-time, on-device multi-output extraction network that generates separate streams for each selected class, achieving robust performance for upto 5 overlapping target sounds and letting users mix their environment by customizing per-class volumes, much like an audio engineer mixes tracks. We optimize the model architecture for multiple compute-limited platforms and demonstrate real-time performance on 6 ms streaming audio chunks. Across real-world environments in previously unseen indoor and outdoor scenarios, our system enables expressive per-class sound control and achieves substantial improvements in target-class enhancement and interference suppression. Our results show that the world need not be heard as a single, undifferentiated stream: with Aurchestra, the soundscape becomes truly programmable.  \nCCS Concepts  \n• Computing methodologies → Artificial intelligence; Machine learning; • Human-centered computing → Ubiquitous  \nThis work is licensed under a Creative Commons Attribution 4 .0 International License. MobiSys ’26, Cambridge, United Kingdom  \n© 2026 Copyright held by the owner/author(s) .  \nACM ISBN 979-8-4007-2027-7/2026/06  \n[https://doi.org/10.1145/3745756.3809210](https://doi.org/10.1145/3745756.3809210)  \nand mobile computing; • Computer systems organization → Embedded and cyber-physical systems.  \nKeywords  \nAugmented hearing, hearables, soundscape control, target sound extraction, sound event detection, on-device machine learning  \nACM Reference Format:  \nSeunghyun Oh,1 Malek Itani,1,2 Aseem Gauri,1 Shyamnath Gollakota1,2 .  \n2026. Fine-grained Soundscape Control for Augmented Hearing. In The 24th Annual International Conference on Mobile Systems, Applications and Services (MobiSys ’26), June 21–25, 2026, Cambridge, United Kingdom. ACM, New York, NY, USA, 21 pages. [https://doi.org/10.1145/3745756.3809210](https://doi.org/10.1145/3745756.3809210)  \n1 Introduction  \nOur acoustic environments are rich, dynamic, and often overwhelming [11]. At any moment, a listener may want to tune in to a nearby sound, amplify an important cue such as an approaching vehicle, dampen distracting chatter, or simply enjoy the surrounding ambience. Yet today’s hearables, both commercial devices and research prototypes, offer only blunt controls: global noise-cancellation modes or a single target-sound focus [55, 56]. In practice, this means users can either pick one sound or suppress all sounds. But the real world is not a toggle switch; it is an orchestra of sounds.  \nIn this paper, we ask an intriguing question: what if users c","cbCaipgIPVAZtrIY","https://ap.wps.com/l/cbCaipgIPVAZtrIY","pdf",18668086,3,1,21,"English","en",105,"# Introduction\n## Motivation and Problem\n## Vision for Expressive Soundscape Control\n## Limitations of Prior Work\n## Research Questions","[{\"question\":\"What problem does Aurchestra address in current hearables?\",\"answer\":\"Current hearables mainly offer global noise suppression or a single-target focus. This prevents users from independently adjusting multiple simultaneous sound sources in real acoustic scenes.\"},{\"question\":\"How does Aurchestra enable fine-grained soundscape control?\",\"answer\":\"It uses (1) a dynamic interface that surfaces only active sound classes and (2) an on-device multi-output extraction network that generates separate streams for each selected class, allowing per-class volume mixing.\"},{\"question\":\"What performance capabilities does the system demonstrate?\",\"answer\":\"The approach targets resource-constrained devices and achieves real-time processing using 6 ms streaming audio chunks, with robust results for up to five overlapping target sound classes.\"}]",1784174671,53,{"code":4,"msg":31,"data":32},"ok",{"site_id":25,"language":24,"slug":33,"title":13,"keywords":34,"description":14,"schema_data":35,"social_meta":86,"head_meta":88,"extra_data":90,"updated_unix":28},"fine-grained-soundscape-control-for-augmented-hearing","",{"@graph":36,"@context":85},[37,53,68],{"@type":38,"itemListElement":39},"BreadcrumbList",[40,44,48,50],{"item":41,"name":42,"@type":43,"position":21},"https://docshare.wps.com","Home","ListItem",{"item":45,"name":46,"@type":43,"position":47},"https://docshare.wps.com/document/","Document",2,{"item":49,"name":12,"@type":43,"position":20},"https://docshare.wps.com/document/research-report/",{"item":51,"name":13,"@type":43,"position":52},"https://docshare.wps.com/document/fine-grained-soundscape-control-for-augmented-hearing/81602/",4,{"url":51,"name":13,"@type":54,"author":55,"headline":13,"publisher":57,"fileFormat":60,"inLanguage":24,"description":14,"dateModified":61,"datePublished":62,"encodingFormat":60,"isAccessibleForFree":63,"interactionStatistic":64},"DigitalDocument",{"name":9,"@type":56},"Person",{"url":41,"name":58,"@type":59},"DocShare","Organization","application/pdf","2026-07-25","2026-07-16",true,{"@type":65,"interactionType":66,"userInteractionCount":20},"InteractionCounter",{"@type":67},"ViewAction",{"@type":69,"mainEntity":70},"FAQPage",[71,77,81],{"name":72,"@type":73,"acceptedAnswer":74},"What problem does Aurchestra address in current hearables?","Question",{"text":75,"@type":76},"Current hearables mainly offer global noise suppression or a single-target focus. This prevents users from independently adjusting multiple simultaneous sound sources in real acoustic scenes.","Answer",{"name":78,"@type":73,"acceptedAnswer":79},"How does Aurchestra enable fine-grained soundscape control?",{"text":80,"@type":76},"It uses (1) a dynamic interface that surfaces only active sound classes and (2) an on-device multi-output extraction network that generates separate streams for each selected class, allowing per-class volume mixing.",{"name":82,"@type":73,"acceptedAnswer":83},"What performance capabilities does the system demonstrate?",{"text":84,"@type":76},"The approach targets resource-constrained devices and achieves real-time processing using 6 ms streaming audio chunks, with robust results for up to five overlapping target sound classes.","https://schema.org",{"og:url":51,"og:type":87,"og:title":13,"og:site_name":58,"og:description":14},"article",{"robots":89,"canonical":51},"index,follow",{"doc_id":7,"site_id":25},{"code":4,"msg":5,"data":92},[93,97,101,105,110,115,120,123,128,131,135],{"id":21,"doc_module":4,"doc_module_name":46,"category_name":94,"show_sort_weight":95,"slug":96},"Story & Novel",90,"story-novel",{"id":47,"doc_module":4,"doc_module_name":46,"category_name":98,"show_sort_weight":99,"slug":100},"Literature",80,"literature",{"id":52,"doc_module":4,"doc_module_name":46,"category_name":102,"show_sort_weight":103,"slug":104},"Exam",70,"exam",{"id":106,"doc_module":4,"doc_module_name":46,"category_name":107,"show_sort_weight":108,"slug":109},5,"Comic",60,"comic",{"id":111,"doc_module":4,"doc_module_name":46,"category_name":112,"show_sort_weight":113,"slug":114},6,"Technology",50,"technology",{"id":116,"doc_module":4,"doc_module_name":46,"category_name":117,"show_sort_weight":118,"slug":119},7,"Healthcare",40,"healthcare",{"id":11,"doc_module":4,"doc_module_name":46,"category_name":12,"show_sort_weight":121,"slug":122},30,"research-report",{"id":124,"doc_module":4,"doc_module_name":46,"category_name":125,"show_sort_weight":126,"slug":127},9,"Religion & Spirituality",20,"religion-spirituality",{"id":126,"doc_module":4,"doc_module_name":46,"category_name":129,"show_sort_weight":126,"slug":130},"World Cup","world-cup",{"id":132,"doc_module":4,"doc_module_name":46,"category_name":133,"show_sort_weight":132,"slug":134},10,"Lifestyle","lifestyle",{"id":136,"doc_module":4,"doc_module_name":46,"category_name":137,"show_sort_weight":106,"slug":138},19,"General","general"]