VLDB 2026 Research / reviewers in the wild / expert
Sameh W. Asaad
dblp:22/3168
· DBLP profile ↗
18ranked-venue papers
1as first author
3since 2021 · last 2024
0000-0001-5652-9892ORCID · reported
Domains — the database's venue-derived domains; a paper can count in several
Systems, architecture and hardware · 15 · 1 first-author · 1 since 2021Applied, interdisciplinary, general and emerging computing · 2 · 1 since 2021Artificial intelligence and machine learning · 1Software engineering, systems software and programming languages · 1 · 1 since 2021
Expertise — from the expertise taxonomy: the topics of the expert's papers under the CCF categories. A weight counts papers with recency: 1 for a paper about the topic, 0.3 when the topic is its context, halved every five years.
| Computer architecture, parallel and distributed computing, and storage systems
6 papers |
Electronic design automation · 30% Cloud and datacenter computing · 16% Storage systems · 16% | |
| Computer networks
2 papers |
Datacenter networks · 56% Software-defined and programmable networks · 44% | |
| Network and information security
1 paper |
Network security · 100% |
Topics — the 17 heaviest of 21, each with the papers that count most for it
| Topic | Weight | Papers | Last | Evidence papers |
|---|---|---|---|---|
Datacenter networks
network-storage co-design |
0.7 | 1 | 2023 | RackBlox: A Software-Defined Rack-Scale Storage System with Network-Storage Co-Design · SOSP 2023 |
Software-defined and programmable networks
programmable data plane |
0.7 | 1 | 2023 | Janus: An Experimental Reconfigurable SmartNIC with P4 Programmability and SDN Isolation · FPGA 2023 |
Cloud and datacenter computing
cloud infrastructure |
0.7 | 1 | 2023 | Janus: An Experimental Reconfigurable SmartNIC with P4 Programmability and SDN Isolation · FPGA 2023 |
Electronic design automation
hardware/software co-design |
0.7 | 1 | 2023 | Janus: An Experimental Reconfigurable SmartNIC with P4 Programmability and SDN Isolation · FPGA 2023 |
Storage systems › flash and SSD › SSD architecture
software-defined flash |
0.7 | 1 | 2023 | RackBlox: A Software-Defined Rack-Scale Storage System with Network-Storage Co-Design · SOSP 2023 |
Reconfigurable computing and FPGAs
FPGA prototyping |
0.6 | 3 | 2017 | Contutto: a novel FPGA-based prototyping platform enabling innovation in the memory subsystem of a server class processor · MICRO 2017 Efficient in-system RTL verification and debugging using FPGAs (abstract only) · FPGA 2012 A cycle-accurate, cycle-reproducible multi-FPGA system for accelerating multi-core processor simulation · FPGA 2012 |
Electronic design automation › hardware verification and test
hardware verification |
0.3 | 2 | 2012 | Efficient in-system RTL verification and debugging using FPGAs (abstract only) · FPGA 2012 A cycle-accurate, cycle-reproducible multi-FPGA system for accelerating multi-core processor simulation · FPGA 2012 |
Emerging computing paradigms
neuromorphic computing |
0.2 | 1 | 2014 | Real-Time Scalable Cortical Computing at 46 Giga-Synaptic OPS/Watt with ~100× Speedup in Time-to-Solution and ~100, 000× Reduction in Energy-to-Solution · SC 2014 |
Electronic design automation › hardware verification and test › functional verification
logic verification |
0.1 | 1 | 2012 | A cycle-accurate, cycle-reproducible multi-FPGA system for accelerating multi-core processor simulation · FPGA 2012 |
Electronic design automation › hardware verification and test › design validation
processor design validation |
0.1 | 1 | 2012 | Efficient in-system RTL verification and debugging using FPGAs (abstract only) · FPGA 2012 |
Memory systems
emerging memory technologies |
0.1 | 1 | 2017 | Contutto: a novel FPGA-based prototyping platform enabling innovation in the memory subsystem of a server class processor · MICRO 2017 |
Memory systems
non-volatile memory |
0.1 | 1 | 2017 | Contutto: a novel FPGA-based prototyping platform enabling innovation in the memory subsystem of a server class processor · MICRO 2017 |
Memory systems › non-volatile memory › non-volatile main memory
NVDIMM |
0.1 | 1 | 2017 | Contutto: a novel FPGA-based prototyping platform enabling innovation in the memory subsystem of a server class processor · MICRO 2017 |
Memory systems › non-volatile memory › magnetic random access memory
STT-MRAM |
0.1 | 1 | 2017 | Contutto: a novel FPGA-based prototyping platform enabling innovation in the memory subsystem of a server class processor · MICRO 2017 |
Hardware accelerators and domain-specific architectures › neural network hardware
brain-inspired computing accelerator |
0.1 | 1 | 2014 | Real-Time Scalable Cortical Computing at 46 Giga-Synaptic OPS/Watt with ~100× Speedup in Time-to-Solution and ~100, 000× Reduction in Energy-to-Solution · SC 2014 |
Energy-efficient computing
power management |
0.1 | 1 | 2014 | Real-Time Scalable Cortical Computing at 46 Giga-Synaptic OPS/Watt with ~100× Speedup in Time-to-Solution and ~100, 000× Reduction in Energy-to-Solution · SC 2014 |
Processor architecture and microarchitecture
chip multiprocessor |
0.0 | 1 | 2012 | A cycle-accurate, cycle-reproducible multi-FPGA system for accelerating multi-core processor simulation · FPGA 2012 |
Methods — techniques the papers use, named apart from their topics
p4 · 2.0hardware-enforced isolation · 2.0hardware offloading · 2.0FPGA prototyping · 0.3event-driven kernel · 0.2chip tiling · 0.2in-system RTL debugging · 0.1design partitioning · 0.1clock synchronization · 0.1FPGA emulation · 0.1
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2024 | UniNet: Accelerating the Container Network Data Plane in IaaS CloudsabstractKubernetes ($K$8s) is a container orchestration plat-form for cloud-based IaaS environments. While it operates on either bare-metal servers or VMs, users prefer VMs for cost savings and agility reasons despite the added network overhead. This overhead, stemming from dual network tunneling at the VM and container levels, degrades performance. To address this, we present UniNet, a SmartNIC-based solution that offloads container-level network tunneling. We designed UniNet to be compatible with leading Container Network Interfaces (CNIs). This approach involves three key elements: (1) transforming VF-based NICs into a container network gateway, (2) offloading the critical path of the data plane functionalities to SmartNICs for enhanced performance and reduced latency, and (3) instituting an isolated control plane that separates VM- and container-level rule insertions, making it tenant-accessible. UniNet boosts CNI throughput by 7.08 x on average, cuts tail latency by 41.6 %, and reduces CPU usage by up to 5.6 x for the receiver and 4.02 x for the sender, respectively. William Gropp, Hubertus Franke, Bharat Sukhwani, Sameh W. Asaad, Jinjun Xiong, Volodymyr V. Kindratenko, Deming Chen |
CLOUD | 6 |
| 2023 | Janus: An Experimental Reconfigurable SmartNIC with P4 Programmability and SDN IsolationabstractDisparate deployment models of cloud computing pose varying requirements on cloud infrastructure components such as networking, storage, provisioning, and security. Infrastructure providers need to study these and often create custom infrastructure components to satisfy these requirements. A major challenge in the research and development of these cloud infrastructure solutions, however, is the availability of customizable platforms for experimentation and trade-off analysis of the various hardware and software components. Most platforms are either general purpose or bespoke solutions created to assist a particular task, too rigid to allow meaningful customization. In this work, we present a 100G reconfigurable smartNIC prototyping platform called Janus that enables cloud infrastructure research and hardware-software co-design of infrastructure components such as hypervisor, secure boot, software defined networking and distributed storage. The platform provides a path to optimize the stack by offloading the functionalities from the host x86 to the embedded processor on the smartNIC and optimize performance by moving pieces to hardware using P4. Further, our platform provides hardware-enforced isolation of cloud network control plane, thereby securing the control plane from the tenants even for bare-metal deployments. Bharat Sukhwani, Mohit Kapur, Alda Ohmacht, Liran Schour, Martin Ohmacht, Chris Ward, Chuck Haymes, Sameh W. Asaad |
FPGA | 8 |
| 2023 | RackBlox: A Software-Defined Rack-Scale Storage System with Network-Storage Co-DesignabstractSoftware-defined networking (SDN) and software-defined flash (SDF) have been serving as the backbone of modern data centers. They are managed separately to handle I/O requests. At first glance, this is a reasonable design by following the rack-scale hierarchical design principles. However, it suffers from suboptimal end-to-end performance, due to the lack of coordination between SDN and SDF. Benjamin Reidys, Yuqi Xue, Daixuan Li, Bharat Sukhwani, Wen-Mei W. Hwu, Deming Chen, Sameh W. Asaad, Jian Huang 0006 |
SOSP | 7 |
| 2017 | Contutto: a novel FPGA-based prototyping platform enabling innovation in the memory subsystem of a server class processorabstractWe demonstrate the use of an FPGA as a memory buffer in a POWER8® system, creating a novel prototyping platform that enables innovation in the memory subsystem of POWER-based servers. Our platform, called ConTutto, is pin-compatible with POWER8 buffered memory DIMMs and plugs into a memory slot of a standard POWER8 processor system, running at aggregate memory channel speeds of 35 GB/s per link. ConTutto, which means "with everything", is a platform to experiment with different memory technologies, such as STT-MRAM and NAND Flash, in an end-to-end system context. Enablement of STT-MRAM and NVDIMM using ConTutto shows up to 12.5x lower latency and 7.5x higher bandwidth compared to the respective technologies when attached to the PCIe bus. Moreover, due to the unique attach-point of the FPGA between the processor and system memory, ConTutto provides a means for in-line acceleration of certain computations on-route to memory, and enables sensitivity analysis for memory latency while running real applications. To the best of our knowledge, ConTutto is the first ever FPGA platform on the memory bus of a server class processor. Bharat Sukhwani, Chuck Haymes, Kyu-Hyoun Kim, Adam J. McPadden, Daniel M. Dreps, Dean Sanner, Jan van Lunteren, Sameh W. Asaad |
MICRO | 9 |
| 2016 | Spatial Predicates Evaluation in the Geohash Domain Using Reconfigurable HardwareabstractAs location sensing devices are becoming ubiquitous, overwhelming amounts of data are being produced by the Internet-of-Things-That-Move. Though analyzing this data presents significant business opportunities, new techniques are needed to attain adequate levels of processing performance. One example is the recently introduced geohash geographical coordinate system that is mainly used for indexing. While geohash codes provide useful inherent properties such as hierarchical and variable-precision coding, traditional spatial algorithms operate on data represented using the conventional latitude/longitude geographical coordinate system, and as such do not take advantage of geohash coding. This paper tackles the evaluation of spatial predicates on geometries defined in the geohash domain, as an alternative to the standard Dimensionally Extended Nine-Intersection Model (DE-9IM). We present the first hardware architecture to efficiently evaluate "contain" and "touch" (internal, external, corner) relations between streams of pairs of geohash codes, in a high throughput (no stall) fashion. Employing FPGAs for exploiting the bit-level granularity of geohash codes, experimental results show (end-to-end) speedup of more than 20× and 90× over highly optimized single-threaded DE-9IM implementations of the contain and touch predicates, respectively. Furthermore, the PCIe-bound FPGA-based solution outperforms a geohash-based multithreaded CPU implementation by ≈1.8× (touch predicate) while using minimal FPGA resources. Dajung Lee, Roger Moussalli, Sameh W. Asaad, Mudhakar Srivatsa |
FCCM | 3 |
| 2015 | Fast and Flexible Conversion of Geohash Codes to and from Latitude/Longitude CoordinatesabstractInsights extracted from spatial queries in geodatabase systems introduce significant opportunities for business intelligence. However, geodatabases are unable to keep up with the required performance due to the massive (and sky-rocketing) amounts of data generated from embedded location-enabled devices. In this paper, we focus on geographic information systems that make use of geohash, specifically, we tackle the kernel of converting geohash codes to and from longitude/latitude pairs. We present the first hardware implementation of a geohash conversion engine operating at wire speed. The presented geohash converter is further enhanced with runtime flexibility with respect to characteristics of the data it can process, furthermore, the architecture allows the user to compromise on performance when limited by hardware resources (design time flexibility). Experimental results of the geohash conversion engine on a Xilinx XC7K325T FPGA show >13X (end-to-end) speedup compared to optimized industry-grade software running on 16 CPU hardware threads. Roger Moussalli, Mudhakar Srivatsa, Sameh W. Asaad |
FCCM | 3 |
| 2014 | Real-Time Scalable Cortical Computing at 46 Giga-Synaptic OPS/Watt with ~100× Speedup in Time-to-Solution and ~100, 000× Reduction in Energy-to-SolutionabstractDrawing on neuroscience, we have developed a parallel, event-driven kernel for neurosynaptic computation, that is efficient with respect to computation, memory, and communication. Building on the previously demonstrated highly optimized software expression of the kernel, here, we demonstrate True North, a co-designed silicon expression of the kernel. True North achieves five orders of magnitude reduction in energy to-solution and two orders of magnitude speedup in time-to solution, when running computer vision applications and complex recurrent neural network simulations. Breaking path with the von Neumann architecture, True North is a 4,096 core, 1 million neuron, and 256 million synapse brain-inspired neurosynaptic processor, that consumes 65mW of power running at real-time and delivers performance of 46 Giga-Synaptic OPS/Watt. We demonstrate seamless tiling of True North chips into arrays, forming a foundation for cortex-like scalability. True North's unprecedented time-to-solution, energy-to-solution, size, scalability, and performance combined with the underlying flexibility of the kernel enable a broad range of cognitive applications. Andrew S. Cassidy, Rodrigo Alvarez-Icaza, Filipp Akopyan, Jun Sawada, John V. Arthur, Paul Merolla, Pallab Datta, Marc González 0001, Brian Taba, Alexander Andreopoulos, Arnon Amir, Steven K. Esser, Jeffrey A. Kusnitz, Rathinakumar Appuswamy, Chuck Haymes, Bernard Brezzo, Roger Moussalli, Ralph Bellofatto, Christian W. Baks, Michael Mastro, Kai Schleupen, Charles E. Cox, Ken Inoue, Steven E. Millman, Nabil Imam, Emmett McQuinn, Yutaka Y. Nakamura, Ivan Vo, Chen Guok, Don Nguyen, Scott Lekuch, Sameh W. Asaad, Daniel J. Friedman, Bryan L. Jackson, Myron Flickner, William P. Risk, Rajit Manohar, Dharmendra S. Modha |
SC | 32 |
| 2014 | A Fully Pipelined FPGA Architecture of a Factored Restricted Boltzmann Machine Artificial Neural NetworkabstractArtificial neural networks (ANNs) are a natural target for hardware acceleration by FPGAs and GPGPUs because commercial-scale applications can require days to weeks to train using CPUs, and the algorithms are highly parallelizable. Previous work on FPGAs has shown how hardware parallelism can be used to accelerate a “Restricted Boltzmann Machine” (RBM) ANN algorithm, and how to distribute computation across multiple FPGAs. Here we describe a fully pipelined parallel architecture that exploits “mini-batch” training (combining many input cases to compute each set of weight updates) to further accelerate ANN training. We implement on an FPGA, for the first time to our knowledge, a more powerful variant of the basic RBM, the “Factored RBM” (fRBM). The fRBM has proved valuable in learning transformations and in discovering features that are present across multiple types of input. We obtain (in simulation) a 100-fold acceleration (vs. CPU software) for an fRBM having N = 256 units in each of its four groups (two input, one output, one intermediate group of units) running on a Virtex-6 LX760 FPGA. Many of the architectural features we implement are applicable not only to fRBMs, but to basic RBMs and other ANN algorithms more broadly. Sameh W. Asaad, Ralph Linsker |
ACM Trans. Reconfigurable Technol. Syst. | 2 |
| 2013 | Accelerating Join Operation for Relational Databases with FPGAsabstractIn this paper, we investigate the use of field programmable gate arrays (FPGAs) to accelerate relational joins. Relational join is one of the most CPU-intensive, yet commonly used, database operations. Hashing can be used to reduce the time complexity from quadratic (naïve) to linear time. However, doing so can introduce false positives to the results which must be resolved. We present a hash-join engine on FPGA that performs hashing, conflict resolution, and joining on a PCIe-attached system, achieving greater than 11x speedup over software. Robert J. Halstead, Bharat Sukhwani, Hong Min, Mathew Thoennes, Parijat Dube, Sameh W. Asaad, Balakrishna Iyer |
FCCM | 6 |
| 2013 | Large Payload Streaming Database Sort and Projection on FPGAsabstractIn recent years, real-time analytics has seen widespread adoption in the business world. While it provides useful business insights and improved market responsiveness, it also adds a computational burden to traditional online transaction processing (OLTP) systems. Analytics queries involve complex database operations such as sort, aggregation, and join that consume significant computational resources, and, when executed on the same system, may affect the performance of OLTP queries. In this paper, we try to address this issue by accelerating two such database operations, namely, projection and sort, using a field programmable gate array (FPGA). Our prototype is implemented on an Alter a Stratix V FPGA and achieves an order of magnitude speedup in the sort operation compared to baseline software. Furthermore, our prototype implements projection in parallel with other query operations on FPGA, thus completely eliminating the cost of projection without consuming any extra cycles on the FPGA. FPGA accelerated sort and projection have been integrated with our previous work on accelerating other query operations [1], making our analytics acceleration prototype on FPGA applicable to a wider variety of queries. Bharat Sukhwani, Mathew Thoennes, Hong Min, Parijat Dube, Bernard Brezzo, Sameh W. Asaad, Donna Dillenberger |
SBAC-PAD | 6 |
| 2012 | Database analytics acceleration using FPGAsabstractBusiness growth and technology advancements have resulted in growing amounts of enterprise data. To gain valuable business insight and competitive advantage, businesses demand the capability of performing real-time analytics on such data. This, however, involves expensive query operations that are very time consuming on traditional CPUs. Additionally, in traditional database management systems (DBMS), the CPU resources are dedicated to mission-critical transactional workloads. Offloading expensive analytics query operations to a co-processor can allow efficient execution of analytics workloads in parallel with transactional workloads. Bharat Sukhwani, Hong Min, Mathew Thoennes, Parijat Dube, Balakrishna Iyer, Bernard Brezzo, Donna Dillenberger, Sameh W. Asaad |
PACT | 8 |
| 2012 | A cycle-accurate, cycle-reproducible multi-FPGA system for accelerating multi-core processor simulationabstractSoftware based tools for simulation are not keeping up with the demands for increased chip and system design complexity. In this paper, we describe a cycle-accurate and cycle-reproducible large-scale FPGA platform that is designed from the ground up to accelerate logic verification of the Bluegene/Q compute node ASIC, a multi-processor SOC implemented in IBM's 45 nm SOI CMOS technology. This paper discusses the challenges for constructing such large-scale FPGA platforms, including design partitioning, clocking & synchronization, and debugging support, as well as our approach for addressing these challenges without sacrificing cycle accuracy and cycle reproducibility. The resulting fullchip simulation of the Bluegene/Q compute node ASIC runs at a simulated processor clock speed of 4 MHz, over 100,000 times faster than the logic level software simulation of the same design. The vast increase in simulation speed provides a new capability in the design cycle that proved to be instrumental in logic verification as well as early software development and performance validation for Bluegene/Q. Sameh W. Asaad, Ralph Bellofatto, Bernard Brezzo, Chuck Haymes, Mohit Kapur, Benjamin D. Parker, Proshanta Saha, Todd Takken, José A. Tierno |
FPGA | 1 |
| 2012 | Efficient in-system RTL verification and debugging using FPGAs (abstract only)abstractFPGAs have become indispensible in processor design, bring-up and debug. Traditionally FPGAs have been used in prototyping, allowing end-users to emulate functionality of a specific component of a processor. However, as the complexity of processors grows, another aspect of processor design, RTL verification, has become a prime target for acceleration using FPGAs. Software-only RTL simulation and verification tools are no longer sufficient for many verification tasks as they often incur long execution time penalties. Software simulation time for a basic Linux kernel bring-up on a BlueGene/Q [1] processor, with 16 user PowerPC A2 cores, for example, could easily exceed several years. Proshanta Saha, Chuck Haymes, Ralph Bellofatto, Bernard Brezzo, Mohit Kapur, Sameh W. Asaad |
FPGA | 6 |
| 2011 | High-Throughput, Lossless Data Compresion on FPGAsabstractLoss less compression is often used before writing data to a storage medium or transmitting across a transmission medium. Compression aids by saving storage space or transmission bandwidth, a decompression operation is performed when the data is subsequently read. Though this scheme has clear benefits, the execution time of compression and decompression is critical to its application in real-time systems. Software compression utilities are often slow, leading to degraded system performance. Hardware-based solutions, on the other hand, often drive large resource requirements and are not amenable to supporting future algorithmic changes. In the current article, we present a high-throughput, streaming, loss less compression algorithm and its efficient implementation on FPGAs. The proposed solution provides a peak throughput of 1GB/sec per engine, with a sustained overall measured throughput of 2.66GB/sec on a PCIe-based FPGA board with two compression and two decompression engines. This result represents an overall speedup of 13.6× over reference software implementation. The proposed design is very lean, and, with multiple engines running in parallel, provides a path to potential speedups of up to two orders of magnitude. In the current implementation, the achievable overall throughput is limited only by the available PCIe bus bandwidth. Bharat Sukhwani, Bülent Abali, Bernard Brezzo, Sameh W. Asaad |
FCCM | 4 |
| 2004 | Design methodology for semi custom processor coresabstractWe describe a semi-custom design methodology for embedded processor cores that was prototyped through the development of a low power high performance DSP core. When compared to the standard ASIC design flow, this methodology enables significant improvement in the speed and power; such benefits are obtained without compromising the generality and flexibility that characterizes the ASIC-based design techniques. Our methodology achieves fast turn-around time in the process from RTL description to post-PD timing results, and exhibits stable convergence on timing; these characteristics enable the application of optimizations spanning multiple levels of the design hierarchy. Such optimizations proved to be much more effective than those that focus only on a single design stage. Victor V. Zyuban, Sameh W. Asaad, Thomas W. Fox, Anne-Marie Haen, Daniel Littrell, Jaime H. Moreno |
ACM Great Lakes Symposium on VLSI | 2 |
| 2003 | Reducing instruction fetch energy with backwards branch control information and bufferingabstractMany emerging applications, e.g. in the embedded and DSP space, are often characterized by their loopy nature where a substantial part of the execution time is spent within a few program phases. Loop buffering techniques have been proposed for capturing and processing these loops in small buffers to reduce the processor`s instruction fetch energy. However, these schemes are limited to straight-line or innermost loops and fail to adequately handle complex loops.In this paper, we propose a dynamic loop buffering mechanism that uses backwards branch control information to identify, capture and process complex loop structures. The DLB controller has been fully implemented in VHDL, synthesized and timed with the IBM Booledozer and Einstimer Synthesis tools, and analyzed for power with the Sequence PowerTheater tool. Our experiments show that the DLB approach, on average, results in a factor of 3 reduction in energy consumption compared to a traditional instruction memory design at an area overhead of about 9%. Jude A. Rivers, Sameh W. Asaad, John-David Wellman, Jaime H. Moreno |
ISLPED | 2 |
| 1998 | Designing a Testable System on a ChipabstractA "system on a chip" is described, which integrates 16 Mbits of DRAM, digital logic, SRAM, three PLLs, and a triple video digital-to-analog converter in a 0.5 micron CMOS DRAM process. Application specific integrated circuit (ASIC) techniques are employed, using multiple DRAM macros with built-in self test (BIST), full level-sensitive scan design (LSSD) logic, and externally accessible analog circuitry. Issues regarding functional debugging, DRAM macro isolation and low cost manufacturing test using only a logic tester are described. Stephen V. Kosonocky, Arthur A. Bright, Kevin W. Warren, Ruud A. Haring, Steve Klepner, Sameh W. Asaad, S. Basavaiah, Bob Havreluk, David F. Heidel, Michael Immediato, Keith A. Jenkins, Rajiv V. Joshi, Benjamin D. Parker, T. V. Rajeevakumar, Kevin Stawiasz |
VTS | 6 |
| 1996 | A novel pneumatic rubber actuator for mobile robot basesabstractA new, pneumatic rubber actuator, named bubbler, realizing linear motion has been designed, developed and tested. The bubbler consists of a silicone rubber slab, and its inside is divided into twelve chambers each having a square cross section. If pressure is applied to successive chambers sequentially, the surface of the bubbler produces a linear traveling wave. It can be used for a mobile robot base or for a conveyor. Experimental results are very promising. The walking speed of the prototype, 63 mm/spl times/60 mm/spl times/9 mm in size, was found to be 3.6 mm/s and the load-to-weight ratio in the order of 40:1. A demonstration board using two bubblers works as a mobile robot base realizing forward/backward and steering motions. Koichi Suzumori, Sameh W. Asaad |
IROS | 2 |