Rahul Agarwal

42 papers A* 2A 2B 3Misc 2Journal 17Unranked 16
YearRankTypeTitle / Venue / Authors
2025 J jnl
CoRR
Aaron K. Baughman, Rahul Agarwal, Eduardo Morales, Gozde Akay
2025 J jnl
CoRR
Rahul Agarwal, Mustafa A. Mohamad
2025 A conf
UMAP
Rahul Agarwal, Amit Jaspal, Saurabh Gupta, Omkar Vichare
2025 J jnl
CoRR
Amit Jaspal, Rahul Agarwal
2025 J jnl
CoRR
Aaron K. Baughman, Gozde Akay, Eduardo Morales, Rahul Agarwal, Preetika Srivastava
2024 conf
ICCCNT
Rakshit Harsh, Raghav Patodiya, Rahul Agarwal, Vivek Mehta
2024 A* conf
KDD
Aaron K. Baughman, Eduardo Morales, Rahul Agarwal, Gozde Akay, Rogério Feris, Tony Johnson, Stephen Hammer, Leonid Karlinsky
2024 J jnl
CoRR
Aaron K. Baughman, Stephen Hammer, Rahul Agarwal, Gozde Akay, Eduardo Morales, Tony Johnson, Leonid Karlinsky, Rogério Feris
2024 J jnl
Medical Biol. Eng. Comput.
Adarsh Sinha, Rahul Agarwal, Vinay Kumar, Nitin Garg, Dhruv Singh Pundir, Harsimran Singh, Ritu Rani, Chinmaya Panigrahy
2023 J jnl
Nat. Lang. Process. J.
Basra Jehangir, Saravanan Radhakrishnan, Rahul Agarwal
2022 conf
ISSCC
John J. Wuu, Rahul Agarwal, Michael Ciraula, Carl Dietz, Brett Johnson, Dave Johnson, Russell Schreiber, Raja Swaminathan, Will Walker, Samuel Naffziger
2021 J jnl
Int. J. Artif. Intell. Mach. Learn.
Sanjit Kumar Dash, Satyam Raj, Rahul Agarwal, Jibitesh Mishra
2021 J jnl
IEEE Access
Rahul Agarwal, Narpat Singh Shekhawat, Sandeep Kumar, Anand Nayyar, Basit Qureshi
2019 J jnl
CoRR
Prakhar Ganesh, Saket Dingliwal, Rahul Agarwal
2017 J jnl
IEEE Trans. Pattern Anal. Mach. Intell.
Rahul Agarwal, Zhe Chen, Sridevi V. Sarma
2016 J jnl
Neural Comput.
Rahul Agarwal, Zhe Chen, Fabian Kloosterman, Matthew A. Wilson, Sridevi V. Sarma
2016 J jnl
Database J. Biol. Databases Curation
Rahul Agarwal, Binayak Kumar, Msk Jayadev, Dhwani Raghav, Ashutosh Singh
2016 conf
3DIC
Luke England, Sukeshwar Kannan, Rahul Agarwal, Daniel Smith
2016 Misc conf
CISS
Rahul Agarwal, Zhe Chen, Fabian Kloosterman, Matthew A. Wilson, Sridevi V. Sarma
2015 conf
IRPS
Sukeshwar Kannan, Rahul Agarwal, Arnaud Bousquet, Geetha Aluri, Hui-Shan Chang
2014 conf
EMBC
Rahul Agarwal, Sabato Santaniello, Sridevi V. Sarma
2013 conf
CICLing (1)
Bhasha Agrawal, Rahul Agarwal, Samar Husain, Dipti Misra Sharma
2012 B conf
LREC
Rahul Agarwal, Bharat Ram Ambati, Anil Kumar Singh
2012 J jnl
PLoS Comput. Biol.
Rahul Agarwal, Sridevi V. Sarma
2012 J jnl
J. Comput. Neurosci.
Rahul Agarwal, Sridevi V. Sarma
2012 conf
ICINCO (1)
Rahul Agarwal, Sridevi V. Sarma
2011 conf
EMBC
Rahul Agarwal, Sridevi V. Sarma
2011 J jnl
IEEE J. Solid State Circuits
Geert Van der Plas, Paresh Limaye, Igor Loi, Abdelkarim Mercha, Herman Oprins, Cristina Torregiani, Steven Thijs, Dimitri Linten, Michele Stucchi, Guruprasad Katti, Dimitrios Velenis, Vladimir Cherman, Bart Vandevelde, Veerle Simons, Ingrid De Wolf, Riet Labie, Dan Perry, Stephane Bronckers, Nikolaos Minas, Miro Cupac, Wouter Ruythooren, Jan Van Olmen, Alain Phommahaxay, Muriel de Potter de ten Broeck, Ann Opdebeeck, Michal Rakowski, Bart De Wachter, Morin Dehan, Marc Nelis, Rahul Agarwal, Antonio Pullini, Federico Angiolini, Luca Benini, Wim Dehaene, Youssef Travaly, Eric Beyne, Paul Marchal
2011 conf
ALR@IJCNLP
Bharat Ram Ambati, Rahul Agarwal, Mridul Gupta, Samar Husain, Dipti Misra Sharma
2010 conf
ISSCC
Geert Van der Plas, Paresh Limaye, Abdelkarim Mercha, Herman Oprins, Cristina Torregiani, Steven Thijs, Dimitri Linten, Michele Stucchi, Guruprasad Katti, Dimitrios Velenis, Domae Shinichi, Vladimir Cherman, Bart Vandevelde, Veerle Simons, Ingrid De Wolf, Riet Labie, Dan Perry, Stephane Bronckers, Nikolaos Minas, Miro Cupac, Wouter Ruythooren, Jan Van Olmen, Alain Phommahaxay, Muriel de Potter de ten Broeck, Ann Opdebeeck, Michal Rakowski, Bart De Wachter, Morin Dehan, Marc Nelis, Rahul Agarwal, Wim Dehaene, Youssef Travaly, Pol Marchal, Eric Beyne
2010 J jnl
IBM J. Res. Dev.
Rahul Agarwal, Saddek Bensalem, Eitan Farchi, Klaus Havelund, Yarden Nir-Buchbinder, Scott D. Stoller, Shmuel Ur, Liqiang Wang
2010 conf
CICC
Geert Van der Plas, Steven Thijs, Dimitri Linten, Guruprasad Katti, Paresh Limaye, Abdelkarim Mercha, Michele Stucchi, Herman Oprins, Bart Vandevelde, Nikolaos Minas, Miro Cupac, Morin Dehan, Marc Nelis, Rahul Agarwal, Wim Dehaene, Youssef Travaly, Eric Beyne, Paul Marchal
2009 conf
3DIC
Yann Civale, Deniz Sabuncuoglu Tezcan, Harold G. G. Philipsen, P. Jaenen, Rahul Agarwal, F. Duval, Philippe Soussan, Youssef Travaly, Eric Beyne
2008 conf
HAPTICS
Mohsen Mahvash, James C. Gwilliam, Rahul Agarwal, Balázs Vágvölgyi, Li-Ming Su, David D. Yuh, Allison M. Okamura
2006 A conf
SIGCSE
Rahul Agarwal, Stephen H. Edwards, Manuel A. Pérez-Quiñones
2006 conf
PADTAD
Rahul Agarwal, Scott D. Stoller
2005 B conf
PPoPP
Amit Sasturkar, Rahul Agarwal, Liqiang Wang, Scott D. Stoller
2005 conf
Haifa Verification Conference
Rahul Agarwal, Liqiang Wang, Scott D. Stoller
2005 Misc conf
ICDCIT
Rahul Agarwal, Mahender Bisht, S. N. Maheshwari, Sanjiva Prasad
2005 conf
ICMENS
Scott Samson, Rahul Agarwal, Sunny Kedia, Weidong Wang, Shinzo Onishi, John Bumgarner
2005 A* conf
ASE
Rahul Agarwal, Amit Sasturkar, Liqiang Wang, Scott D. Stoller
2004 B conf
VMCAI
Rahul Agarwal, Scott D. Stoller
redb/extractors/macho_extractors/macho_segments.py
← Index redb/extractors/macho_extractors/macho_segments.py python
import hashlib
import inspect
import base64
from datetime import datetime, timezone
from typing import Any

from redb.extractors.enum import Tag
from redb.extractors.macho_extractor import MachOExtractor
from redb.models.dataclasses import MachOSegment


class MachOSegmentExtractor(MachOExtractor):

    def __init__(
        self,
        filepath,
        log,
        exporters=None,
        index_prefix=None,
        elastic_index=None,
        known_benign=False,
        known_malicious=False,
        macho=None,
    ):
        super().__init__(
            filepath,
            log,
            exporters,
            index_prefix,
            elastic_index,
            known_benign,
            known_malicious,
            macho,
        )
        self.elastic_index = self.index_prefix + "-macho_segments"
        self.log.debug(inspect.currentframe().f_code.co_name)

    def _is_empty_result(self, extracted_data) -> bool:
        """
        Override: Empty segments is an ERROR, not a valid empty case.
        A valid MachO file must have segments (at minimum __PAGEZERO, __TEXT).
        """
        # Always return False - empty segments should be treated as an error
        return False

    def tag(self):
        return Tag.MACHO_SEGMENT.value

    def _extract_segments_for_arch(self, arch_name):
        """Extract segment information for a specific architecture."""
        self.log.debug(f"Extracting segments for architecture: {arch_name}")
        segments = []

        try:
            # Get segments using new API with architecture parameter
            segments_data = self.macho.get_segments(arch=arch_name)
            if not segments_data:
                return segments

            # Extract segments for this architecture
            for segment in segments_data:
                try:
                    segment_name = segment.get('segname', 'Unknown')

                    # Calculate segment hash
                    segment_data = self._get_segment_data(segment)
                    if segment_data:
                        seg_sha256 = hashlib.sha256(segment_data).hexdigest()
                    else:
                        seg_sha256 = ""

                    # Use entropy already calculated by machofile module, rounded to 3 decimal places
                    seg_entropy = round(segment.get('entropy', 0.0), 3)

                    # Create segment dataclass with architecture info
                    macho_segment = MachOSegment(
                        segment_name=segment_name,
                        segment_vaddr=segment.get('vaddr', 0),
                        segment_vsize=segment.get('vsize', 0),
                        segment_offset=segment.get('offset', 0),
                        segment_size=segment.get('size', 0),
                        segment_max_vm_protection=segment.get('max_vm_protection', 0),
                        segment_initial_vm_protection=segment.get('initial_vm_protection', 0),
                        segment_nsects=segment.get('nsects', 0),
                        segment_flags=segment.get('flags', 0),
                        segment_entropy=seg_entropy,
                        segment_sha256=seg_sha256,
                    )
                    # Add architecture info to the segment
                    macho_segment.architecture = arch_name
                    segments.append(macho_segment)

                except Exception as e:
                    self.log.warning(
                        f'Unable to process segment "{segment.get("segname", "Unknown")}" for architecture {arch_name} in {self.hash.sha256}: {e}'
                    )
                    continue

            return segments

        except Exception as e:
            self.log.error(f"Error extracting MachO segments for architecture {arch_name}: {e}")
            return segments

    def _extract_segments(self):
        """Extract segment information from all architectures in the MachO binary."""
        self.log.debug(inspect.currentframe().f_code.co_name)
        segments = []

        if not self.macho:
            return segments

        try:
            # Get architectures using new API (already parsed in base class)
            architectures = self.macho.get_architectures()
            if not architectures:
                return segments

            # Process each architecture
            for arch_name in architectures:
                arch_segments = self._extract_segments_for_arch(arch_name)
                segments.extend(arch_segments)

            return segments

        except Exception as e:
            self.log.error(f"Error extracting MachO segments: {e}")
            return segments

    def _get_segment_data(self, segment):
        """Get the raw data for a segment."""
        try:
            offset = segment.get('offset', 0)
            size = segment.get('size', 0)
            
            if size == 0:
                return None
            
            # Read segment data from file
            with open(self.filepath, 'rb') as f:
                f.seek(offset)
                return f.read(size)
                
        except Exception as e:
            self.log.warning(f"Error reading segment data: {e}")
            return None

    def extract(self):
        self.log.debug(inspect.currentframe().f_code.co_name)
        try:
            segments = self._extract_segments()
            return segments
        except Exception as e:
            self.log.error(f"Error extracting MachO segments: {e}")
            return None

    def prepare_export_data(self, exporter_type: str) -> Any:
        if exporter_type == "ElasticsearchExporter":
            return self.extract()
        elif exporter_type == "ClickHouseExporter":
            if not self.macho:
                return None

            # Get architectures (macho is already parsed in base class)
            try:
                architectures = self.macho.get_architectures()
                is_fat = len(architectures) > 1
            except Exception as e:
                self.log.error(f"Could not get architectures: {e}")
                return None

            data = []
            current_time = datetime.now(timezone.utc)

            # Loop through each architecture (1 for single, multiple for FAT)
            for arch_name in architectures:
                # Extract segments for this specific architecture
                segments = self._extract_segments_for_arch(arch_name)
                if not segments:
                    continue

                # Get architecture-specific sha256 and header info
                try:
                    arch_general_info = self.macho.get_general_info(arch=arch_name)
                    arch_sha256 = arch_general_info.get('SHA256', self.sha256)

                    # Get raw architecture value
                    arch_header_raw = self.macho.get_macho_header(arch=arch_name)
                    arch_cputype_raw = arch_header_raw.get('cputype', 0) if arch_header_raw else 0
                except Exception as e:
                    self.log.warning(f"Could not get arch-specific data for {arch_name}: {e}")
                    arch_sha256 = self.sha256
                    arch_cputype_raw = 0

                for segment in segments:
                    data.append([
                        arch_sha256,                          # sha256 (architecture-specific)
                        segment.segment_name,                 # segment_name
                        segment.segment_vaddr,                # segment_vaddr
                        segment.segment_vsize,                # segment_vsize
                        segment.segment_offset,               # segment_offset
                        segment.segment_size,                 # segment_size
                        segment.segment_max_vm_protection,    # segment_max_vm_protection
                        segment.segment_initial_vm_protection, # segment_initial_vm_protection
                        segment.segment_nsects,               # segment_nsects
                        segment.segment_flags,                # segment_flags
                        segment.segment_entropy,              # segment_entropy
                        segment.segment_sha256,               # segment_sha256
                        current_time,                         # analysis_date
                    ])

            column_names = [
                'sha256',
                'segment_name', 'segment_vaddr', 'segment_vsize', 'segment_offset',
                'segment_size', 'segment_max_vm_protection', 'segment_initial_vm_protection',
                'segment_nsects', 'segment_flags', 'segment_entropy', 'segment_sha256',
                'analysis_date'
            ]

            if not data:
                return None

            column_type_names = [
                'FixedString(64)',
                'String', 'UInt64', 'UInt64', 'UInt64',
                'UInt64', 'UInt32', 'UInt32',
                'UInt32', 'UInt32', 'Float64', 'FixedString(64)',
                'DateTime64(3, \'UTC\')'
            ]

            return (data, column_names, column_type_names)

        return None

    def get_clickhouse_table(self) -> str:
        return "redb_macho_segments"