Rahul Garg

55 papers A* 5A 7B 4Misc 5Journal 14Unranked 20
YearRankTypeTitle / Venue / Authors
2024 J jnl
CoRR
Shivansh Chandra Tripathi, Rahul Garg
2024 conf
COMAD/CODS
Amit Prasad, Rahul Garg, Yogish Sabharwal
2024 J jnl
CoRR
Shivansh Chandra Tripathi, Rahul Garg
2023 conf
PReMI
Shivansh Chandra Tripathi, Rahul Garg
2022 J jnl
CoRR
Mustafa Chasmai, Nirjhar Das, Aman Bhardwaj, Rahul Garg
2022 J jnl
SN Comput. Sci.
Mustafa Chasmai, Nirjhar Das, Aman Bhardwaj, Rahul Garg
2022 conf
ICHI
Raunak Jain, Mrityunjai Singh, A. Ravishankar Rao, Rahul Garg
2022 J jnl
NeuroImage
Vaibhav Tripathi, Rahul Garg
2020 conf
COMAD/CODS
Abhishek Goyal, Rahul Garg
2019 Misc conf
HiPC
Shreenivas Bharadwaj Venkataramanan, Rahul Garg, Yogish Sabharwal
2019 conf
COMAD/CODS
Aashish Nagpal, Chayan Sharma, Rahul Garg, Pawan Kumar
2018 B conf
LREC
Shubham Bhardwaj, Neelamadhav Gantayat, Nikhil Chaturvedi, Rahul Garg, Sumeet Agarwal
2016 J jnl
ACM Trans. Algorithms
Lisa Fleischer, Rahul Garg, Sanjiv Kapoor, Rohit Khandekar, Amin Saberi
2011 conf
ISBI
A. Ravishankar Rao, Rahul Garg, Guillermo A. Cecchi
2011 A conf
AISTATS
Rahul Garg, Rohit Khandekar
2011 conf
ISBI
Rahul Garg, Guillermo A. Cecchi, A. Ravishankar Rao
2011 J jnl
NeuroImage
Rahul Garg, Guillermo A. Cecchi, A. Ravishankar Rao
2010 Misc conf
HiPC
Rahul Garg, Perwez Shahabuddin, Akshat Verma
2010 J jnl
IEEE Trans. Parallel Distributed Syst.
Rahul Garg, Vijay K. Garg, Yogish Sabharwal
2009 conf
MICCAI (1)
Guillermo A. Cecchi, Rahul Garg, A. Ravishankar Rao
2009 A* conf
ICML
Rahul Garg, Rohit Khandekar
2009 A conf
IPDPS
Vikas Aggarwal, Yogish Sabharwal, Rahul Garg, Philip Heidelberger
2009 J jnl
NeuroImage
Melissa K. Carroll, Guillermo A. Cecchi, Irina Rish, Rahul Garg, A. Ravishankar Rao
2008 conf
WINE
Lisa Fleischer, Rahul Garg, Sanjiv Kapoor, Rohit Khandekar, Amin Saberi
2008 conf
ISBI
Guillermo A. Cecchi, Rahul Garg, A. Ravishankar Rao
2008 B conf
ICPP
Sameer Kumar, Yogish Sabharwal, Rahul Garg, Philip Heidelberger
2008 Misc conf
HiPC
Yogish Sabharwal, Saurabh Kumar Garg, Rahul Garg, John A. Gunnels, Ramendra K. Sahoo
2007 A* conf
ICCV
Manik Varma, Rahul Garg
2007 conf
WINE
Rahul Garg, Sanjiv Kapoor
2007 conf
LCPC
Christopher Barton, Calin Cascaval, George Almási, Rahul Garg, José Nelson Amaral, Montse Farreras
2006 J jnl
Math. Oper. Res.
Rahul Garg, Sanjiv Kapoor
2006 conf
WINE
Pradeep Dubey, Rahul Garg, Bernard De Meyer
2006 conf
WINE
Pradeep Dubey, Rahul Garg
2006 A conf
SC
Hiroshi Akiba, Tomonobu Ohyama, Yoshinoir Shibata, Kiyoshi Yuyama, Yoshikazu Katai, Ryuichi Takeuchi, Takeshi Hoshino, Shinobu Yoshimura, Hirohisa Noguchi, Manish Gupta, John A. Gunnels, Vernon Austel, Yogish Sabharwal, Rahul Garg, Shoji Kato, Takashi Kawakami, Satoru Todokoro, Junko Ikeda
2006 Misc conf
HiPC
Rahul Garg, Pradipta De
2006 A conf
SC
Rahul Garg, Yogish Sabharwal
2006 conf
SIGMETRICS/Performance
Rahul Garg, Yogish Sabharwal
2006 conf
WINE
Rahul Garg, Sanjiv Kapoor
2006 A conf
ICS
Rahul Garg, Vijay K. Garg, Yogish Sabharwal
2005 J jnl
Electron. Commer. Res.
Vipul Bansal, Rahul Garg
2005 Misc conf
HiPC
Saurabh Agarwal, Rahul Garg, Nisheeth K. Vishnoi
2004 A conf
ICS
Saurabh Agarwal, Rahul Garg, Meeta Sharma Gupta, José E. Moreira
2004 conf
APPROX-RANDOM
Rahul Garg, Sanjiv Kapoor, Vijay V. Vazirani
2004 A* conf
STOC
Rahul Garg, Sanjiv Kapoor
2003 A* conf
EC
Rahul Garg, Vijay Kumar, Atri Rudra, Akshat Verma
2002 J jnl
Comput. Commun. Rev.
Rahul Garg, Abhinav Kamra, Varun Khurana
2002 A conf
SC
Narasimha R. Adiga, George Almási, George S. Almási, Yariv Aridor, Rajkishore Barik, Daniel K. Beece, Ralph Bellofatto, Gyan Bhanot, Randy Bickford, Matthias A. Blumrich, Arthur A. Bright, José R. Brunheroto, Calin Cascaval, José G. Castaños, Waiman Chan, Luis Ceze, Paul Coteus, Siddhartha Chatterjee, Dong Chen, George L.-T. Chiu, Thomas M. Cipolla, Paul Crumley, K. M. Desai, Alina Deutsch, Tamar Domany, Marc Boris Dombrowa, Wilm E. Donath, Maria Eleftheriou, C. Christopher Erway, J. Esch, Blake G. Fitch, Joseph Gagliano, Alan Gara, Rahul Garg, Robert S. Germain, Mark Giampapa, Balaji Gopalsamy, John A. Gunnels, Manish Gupta, Fred G. Gustavson, Shawn Hall, Ruud A. Haring, David F. Heidel, Philip Heidelberger, Lorraine Herger, Dirk Hoenicke, R. D. Jackson, T. Jamal-Eddine, Gerard V. Kopcsay, Elie Krevat, Manish P. Kurhekar, Alphonso P. Lanzetta, Derek Lieber, L. K. Liu, M. Lu, Mark P. Mendell, A. Misra, Yosef Moatti, Lawrence S. Mok, José E. Moreira, Ben J. Nathanson, Matthew Newton, Martin Ohmacht, Adam J. Oliner, Vinayaka Pandit, R. B. Pudota, Rick A. Rand, Richard D. Regan, Bradley Rubin, Albert E. Ruehli, Silvius Vasile Rus, Ramendra K. Sahoo, Alda Sanomiya, Eugen Schenfeld, M. Sharma, Edi Shmueli, Sarabjeet Singh, Peilin Song, Vijay Srinivasan, Burkhard D. Steinmacher-Burow, Karin Strauss, Christopher W. Surovic, Richard A. Swetz, Todd Takken, R. Brett Tremaine, Mickey Tsao, Arun R. Umamaheshwaran, P. Verma, Pavlos Vranas, T. J. Christopher Ward, Michael E. Wazlowski, W. Barrett, C. Engel, B. Drehmel, B. Hilgart, D. Hill, F. Kasemkhani, David J. Krolak, Chun-Tao Li, Thomas A. Liebsch, James A. Marcella, A. Muff, A. Okomo, M. Rouse, A. Schram, M. Tubbs, G. Ulsh, Charles D. Wait, J. Wittrup, Myung Bae, Kenneth A. Dockser, Lynn Kissel, Mark K. Seager, Jeffrey S. Vetter, K. Yates
2001 conf
USENIX ATC, General Track
Rahul Garg, Parul A. Mittal, Vikas Agarwal, Natwar Modani
2001 conf
RANDOM-APPROX
Rahul Garg, Vijay Kumar, Vinayaka Pandit
2001 J jnl
SIGecom Exch.
Vipul Bansal, Rahul Garg
2001 B conf
WADS
Amitabha Bagchi, Amitabh Chaudhary, Rahul Garg, Michael T. Goodrich, Vijay Kumar
2000 A* conf
INFOCOM
Rahul Garg, Huzur Saran
1999 conf
Broadband Communications
Rahul Garg, Raphael Rom
1999 J jnl
Comput. Networks
Rahul Garg, Xiaoqiang Chen
1995 B conf
ICIP
Robert J. Safranek, Charles R. Kalmanek Jr., Rahul Garg
redb/extractors/extractor.py
← Index redb/extractors/extractor.py python
import hashlib
import inspect
from abc import ABCMeta, abstractmethod
from dataclasses import asdict
from functools import cached_property
from datetime import datetime, timezone
import math
from typing import Counter, List, Optional, Dict, Any, Tuple
from redb import settings
from redb.models.dataclasses import Hash
from .database_exporters import DatabaseExporter, ElasticsearchExporter, ClickHouseExporter, PrintExporter

class Extractor(metaclass=ABCMeta):
    def __init__(
        self,
        filepath: str,
        log: Any,
        exporters: Optional[List[DatabaseExporter]] = None,
        index_prefix: Optional[str] = None,
        source: Optional[str] = None,
        elastic_index: Optional[str] = None,
        known_benign: bool = False,
        known_malicious: bool = False,
        precomputed_hashes: Optional[Dict[str, str]] = None,
    ):
        self.log = log
        self.log.debug(f"Creating {self.__class__.__name__}")
        self.filepath = filepath
        self.source = source
        self.exporters = exporters or []
        self.index_prefix = index_prefix if index_prefix else settings.ELASTIC_BINARIES_COLLECTION
        self.elastic_index = self.index_prefix + (elastic_index if elastic_index else "")
        self.known_benign = known_benign
        self.known_malicious = known_malicious

        # Use precomputed hashes if provided (e.g., from machofile, pefile)
        # Otherwise compute them from binary
        if precomputed_hashes:
            self.md5 = precomputed_hashes.get('md5') or precomputed_hashes.get('MD5')
            self.sha1 = precomputed_hashes.get('sha1') or precomputed_hashes.get('SHA1')
            self.sha256 = precomputed_hashes.get('sha256') or precomputed_hashes.get('SHA256')
        else:
            self.md5 = hashlib.md5(self.binary).hexdigest()
            self.sha1 = hashlib.sha1(self.binary).hexdigest()
            self.sha256 = hashlib.sha256(self.binary).hexdigest()
        self.hash = Hash(self.md5, self.sha1, self.sha256)

    @cached_property
    def binary(self):
        with open(self.filepath, "rb") as f:
            data = f.read()
        return data

    @property
    @abstractmethod
    def tag(self):
        pass

    @abstractmethod
    def extract(self):
        """
        this method defines the extracted data
        """

    @staticmethod
    def process_binary_string(s):
        # Remove \x00 padding
        s = s.rstrip(b"\x00")

        # Check if there are any non-printable characters
        has_non_printable = any(byte < 32 or byte > 126 for byte in s)

        if not has_non_printable:
            # If all characters are printable, decode the string
            return s.decode()
        else:
            # If there are non-printable characters, represent them as \xDD
            return "".join(
                [
                    f"\\x{byte:02x}" if byte < 32 or byte > 126 else chr(byte)
                    for byte in s
                ]
            )

    @staticmethod
    def remove_non_utf8(binary_string):
        decoded = b""
        for i in range(len(binary_string)):
            try:
                # Try to decode each byte
                char = binary_string[i : i + 1].decode("utf-8")
                decoded += char.encode("utf-8")
            except UnicodeDecodeError:
                # Skip this byte if it can't be decoded
                continue
        return decoded

    def calculate_entropy(self, data):
        """Calculate the entropy of a chunk of data.
        Based on pefile.SectionStructure.entropy_H.
        """
        # self.log.debug(inspect.currentframe().f_code.co_name)
        if not data:
            return 0.0

        if type(data) == str:
            counts = Counter(data)
            frequencies = ((i / len(data)) for i in counts.values())
            return - sum(f * math.log(f, 2) for f in frequencies)
        else:
            occurences = Counter(bytearray(data))
            entropy = 0
            for x in occurences.values():
                p_x = float(x) / len(data)
                entropy -= p_x * math.log(p_x, 2)
            return entropy

    @abstractmethod
    def prepare_export_data(self, exporter_type: str) -> Tuple[List[Any], List[str], List[str]]:
        """
        Prepare data for specific export type
        Returns:
            Tuple containing:
            - data: List of values to insert
            - column_names: List of column names
            - column_type_names: List of column types
        """
        pass

    def export_data(self):
        """Export data to all configured exporters

        Returns:
            True: Export succeeded
            False: Export failed (actual error)
            None: No data to export (not an error, e.g., no overlay, no signature)
        """
        self.log.debug(inspect.currentframe().f_code.co_name)
        success = True
        extracted_data = self.extract()

        if extracted_data is None:
            self.log.debug("extract() returned None, skipping export")
            return None  # No data to export, not a failure
            
        for exporter in self.exporters:
            if isinstance(exporter, PrintExporter):
                # For PrintExporter, we pass the extracted data directly
                success &= exporter.export(extracted_data)
            else:
                # Get the data prepared for this specific exporter type
                export_data = self.prepare_export_data(exporter.__class__.__name__)
                
                if export_data is None:
                    self.log.debug(f"prepare_export_data returned None for {exporter.__class__.__name__}")
                    return False
                
                if isinstance(exporter, ElasticsearchExporter):
                    success &= exporter.export(
                        export_data,
                        index=self.elastic_index,
                        tag=self.tag(),
                        hashes=asdict(self.hash),
                        known_benign=self.known_benign,
                        known_malicious=self.known_malicious
                    )
                elif isinstance(exporter, ClickHouseExporter):
                    # For ClickHouse, we need to pass the table name and the prepared data
                    success &= exporter.export(
                        export_data,
                        table=self.get_clickhouse_table(),
                        # Add these parameters explicitly
                        column_names=export_data[1] if isinstance(export_data, tuple) else None,
                        column_type_names=export_data[2] if isinstance(export_data, tuple) else None
                    )
                
        return success

    @abstractmethod
    def get_clickhouse_table(self) -> str:
        """Return the appropriate ClickHouse table name"""
        pass
    # def export_to_elastic(self, list_of_dataclasses, tag=None):
    #     self.log.debug(inspect.currentframe().f_code.co_name)

    #     if not isinstance(list_of_dataclasses, list):
    #         self.log.error("Called export_to_elastic wrongly")

    #     now_t = datetime.now(timezone.utc).strftime("%Y-%m-%d %H:%M:%S")
    #     for dataclass_ in list_of_dataclasses:
    #         if not settings.ELASTIC_CLIENT.ping():
    #             self.log.error("[CONNECTION ERROR] ping to elastic failed")
            
    #         # Convert dataclass to dict and filter out None values
    #         document = {k: v for k, v in asdict(dataclass_).items() if v is not None}
            
    #         if tag:
    #             document["tag"] = [tag, self.tag()]
    #         else:
    #             document["tag"] = self.tag()
    #         hashes = asdict(self.hash)
    #         document |= hashes
    #         document["timestamp_utc"] = now_t
    #         # document["source"] = self.source
    #         document["known_benign"] = self.known_benign
    #         document["known_malicious"] = self.known_malicious

    #         if "_id" in document:
    #             tmp_id = document.pop("_id") + document["sha256"]
    #             _id = hashlib.sha256(tmp_id.encode()).hexdigest()
    #         else:
    #             _id = document["sha256"]

    #         # self.log.debug(f"[DEBUG] about to export {type(document)} {document}")
    #         try:
    #             doc_dump = json.dumps(document)
    #         except TypeError as e:
    #             self.log.error(
    #                 f"Failed export of document. " f"full document: {document}"
    #             )
    #             raise e

    #         # body={"doc": doc_dump,
    #         #       "doc_as_upsert": True  # Create the document if it doesn't exist
    #         # }

    #         # Check if the index exists, and create it if it doesn't
    #         # if not settings.ELASTIC_CLIENT.indices.exists(index=self.elastic_index):
    #         #     settings.ELASTIC_CLIENT.indices.create(index=self.elastic_index)
    #         # pprint(doc_dump) #DEBUG
    #         # print("[DEBUG] _id: " + _id)
    #         # print("[DEBUG] index: " + self.elastic_index)
    #         settings.ELASTIC_CLIENT.index(
    #             index=self.elastic_index, id=_id, document=doc_dump
    #         )