James Evans

71 papers A* 5A 2B 1Journal 59Unranked 4
YearRankTypeTitle / Venue / Authors
2026 J jnl
CoRR
James Evans, Benjamin Bratton, Blaise Agüera y Arcas
2026 J jnl
ACM Trans. Soc. Comput.
Pu Zhang, Zheng Wei, Muzhi Zhou, James Evans, Pan Hui
2026 J jnl
CoRR
Jeffrey W. Lockhart, Jamshid Sourati, Feng Shi, James Evans
2026 J jnl
CoRR
Jio Oh, Steven Euijong Whang, James Evans, Jindong Wang
2026 J jnl
CoRR
Likun Cao, James Evans
2026 J jnl
CoRR
Laura Ferrarotti, Gian Maria Campedelli, Roberto Dessì, Andrea Baronchelli, Giovanni Iacca, Kathleen M. Carley, Alex Pentland, Joel Z. Leibo, James Evans, Bruno Lepri
2026 J jnl
CoRR
Anirudh Ajith, Amanpreet Singh, Jay DeYoung, Nadav Kunievsky, Austin C. Kozlowski, Oyvind Tafjord, James Evans, Daniel S. Weld, Tom Hope, Doug Downey
2026 J jnl
CoRR
Junsol Kim, Shiyang Lai, Nino Scherrer, Blaise Agüera y Arcas, James Evans
2026 J jnl
CoRR
Carolina Biliotti, Massimo Riccaboni, Jeffrey W. Lockhart, James Evans
2026 J jnl
CoRR
Junsol Kim, Winnie Street, Roberta Rocca, Diane M. Korngiebel, Adam Waytz, James Evans, Geoff Keeling
2026 J jnl
CoRR
Lin Chen, Fengli Xu, Esteban Moro, Pan Hui, Yong Li, James Evans
2026 J jnl
ACM Trans. Soc. Comput.
Xiaoming Fu, James Evans
2025 J jnl
CoRR
Matthew O. Jackson, Qiaozhu Mei, Stephanie W. Wang, Yutong Xie, Walter Yuan, Seth Benzell, Erik Brynjolfsson, Colin F. Camerer, James Evans, Brian Jabarian, Jon M. Kleinberg, Juanjuan Meng, Sendhil Mullainathan, Asuman Ozdaglar, Thomas Pfeiffer, Moshe Tennenholtz, Robb Willer, Diyi Yang, Teng Ye
2025 J jnl
CoRR
Shiyang Lai, Junsol Kim, Nadav Kunievsky, Yujin Potter, James Evans
2025 J jnl
CoRR
Ningzi Li, Shiyang Lai, James Evans
2025 J jnl
CoRR
Jinghua Piao, Zhihong Lu, Chen Gao, Fengli Xu, Fernando P. Santos, Yong Li, James Evans
2025 J jnl
CoRR
Jingjing Qu, Kejia Hu, Jun Zhu, Wenhao Li, Teng Wang, Zhiyun Chen, Yulei Ye, Chaochao Lu, Aimin Zhou, Xiangfeng Wang, James Evans
2025 J jnl
CoRR
Hongbo Fang, James Evans
2025 J jnl
CoRR
Zheng Wei, Mingchen Li, Zeqian Zhang, Ruibin Yuan, Pan Hui, Huamin Qu, James Evans, Maneesh Agrawala, Anyi Rao
2025 J jnl
CoRR
Zhen Zhang, James Evans
2025 A* conf
ICLR
Junsol Kim, James Evans, Aaron Schein
2025 J jnl
CoRR
Junsol Kim, James Evans, Aaron Schein
2025 J jnl
CoRR
Veniamin Veselovsky, Berke Argin, Benedikt Stroebl, Chris Wendler, Robert West, James Evans, Thomas L. Griffiths, Arvind Narayanan
2025 J jnl
CoRR
Zhilun Zhou, Jing Yi Wang, Nicholas Sukiennik, Chen Gao, Fengli Xu, Yong Li, James Evans
2025 J jnl
CoRR
HyunJin Kim, Xiaoyuan Yi, Jing Yao, Muhua Huang, JinYeong Bak, James Evans, Xing Xie
2025 J jnl
CoRR
Likun Cao, Rui Pan, James Evans
2025 J jnl
CoRR
Jiaxi Yang, Mengqi Zhang, Yiqiao Jin, Hao Chen, Qingsong Wen, Lu Lin, Yi He, Weijie Xu, James Evans, Jindong Wang
2025 conf
ACL (3)
Jing Yao, Xiaoyuan Yi, Shitong Duan, Jindong Wang, Yuzhuo Bai, Muhua Huang, Yang Ou, Scarlett Li, Peng Zhang, Tun Lu, Zhicheng Dou, Maosong Sun, James Evans, Xing Xie
2025 J jnl
CoRR
Iason Gabriel, Geoff Keeling, Arianna Manzini, James Evans
2024 A* conf
NeurIPS
Chengxing Xie, Canyu Chen, Feiran Jia, Ziyu Ye, Shiyang Lai, Kai Shu, Jindong Gu, Adel Bibi, Ziniu Hu, David Jurgens, James Evans, Philip Torr, Bernard Ghanem, Guohao Li
2024 J jnl
CoRR
Blaise Agüera y Arcas, Jyrki Alakuijala, James Evans, Ben Laurie, Alexander Mordvintsev, Eyvind Niklasson, Ettore Randazzo, Luca Versari
2024 J jnl
CoRR
Muhua Huang, Xijuan Zhang, Christopher Soto, James Evans
2024 J jnl
CoRR
Shiyang Lai, Yujin Potter, Junsol Kim, Richard Zhuang, Dawn Song, James Evans
2024 A* conf
EMNLP
Yujin Potter, Shiyang Lai, Junsol Kim, James Evans, Dawn Song
2024 J jnl
CoRR
Yujin Potter, Shiyang Lai, Junsol Kim, James Evans, Dawn Song
2024 A* conf
ICML
Shiyang Lai, Yujin Potter, Junsol Kim, Richard Zhuang, Dawn Song, James Evans
2024 J jnl
CoRR
Hongbo Fang, Patrick S. Park, James Evans, James D. Herbsleb, Bogdan Vasilescu
2023 J jnl
CoRR
Jamshid Sourati, James Evans
2023 J jnl
CoRR
Junsol Kim, Zhao Wang, Haohan Shi, Hsin-Keng Ling, James Evans
2023 J jnl
J. Soc. Comput.
James Evans, Xiaoming Fu, Jar-der Luo
2022 B conf
IEEE Big Data
Matthew L. Thomas, Gavin Shaddick, David Topping, Karyn Morrissey, Thomas J. Brannan, Mike Diessner, Ruth C. E. Bowyer, Stefan Siegert, Hugh Coe, James Evans, Fernando Benitez-Paez, James V. Zidek
2022 J jnl
CoRR
Jamshid Sourati, James Evans
2021 J jnl
CoRR
Jamshid Sourati, James Evans
2021 conf
EMNLP (1)
Jeremiah Milbauer, Adarsh Mathew, James Evans
2021 J jnl
CoRR
Eamon Duede, Misha Teplitskiy, Karim R. Lakhani, James Evans
2021 J jnl
Discret. Math.
John Bamberg, James Evans
2021 J jnl
CoRR
Brendan Chambers, James Evans
2021 J jnl
CoRR
David H. Wolpert, Michael H. Price, Stefani A. Crabtree, Timothy A. Kohler, Jürgen Jost, James Evans, Peter F. Stadler, Hajime Shimao, Manfred D. Laubichler
2020 A conf
SIGCSE
Jérémie O. Lumbroso, James Evans
2020 J jnl
J. Soc. Comput.
James Evans
2019 J jnl
CoRR
Timmy Li, Yi Huang, James Evans, Ishanu Chattopadhyay
2019 J jnl
CoRR
Feng Shi, James Evans
2019 J jnl
CoRR
Nandana Sengupta, Nati Srebro, James Evans
2018 A conf
AISTATS
Sumeet Katariya, Lalit K. Jain, Nandana Sengupta, James Evans, Robert Nowak
2018 J jnl
CoRR
Sumeet Katariya, Lalit K. Jain, Nandana Sengupta, James Evans, Robert Nowak
2018 J jnl
CoRR
Shahab Asoodeh, Tingran Gao, James Evans
2018 J jnl
CoRR
Misha Teplitskiy, Daniel E. Acuna, Aida Elamrani-Raoult, Konrad P. Körding, James Evans
2018 J jnl
CoRR
Tingran Gao, Shahab Asoodeh, Yi Huang, James Evans
2017 J jnl
Informatics
Krassimira Paskaleva, James Evans, Christopher Martin, Trond Linjordet, Dujuan Yang, Andrew Karvonen
2017 J jnl
CoRR
Feng Shi, Misha Teplitskiy, Eamon Duede, James Evans
2015 A* conf
EMNLP
Jingwei Zhang, Aaron Gerow, Jaan Altosaar, James Evans, Richard Jean So
2015 J jnl
CoRR
Jingwei Zhang, Aaron Gerow, Jaan Altosaar, James Evans, Richard Jean So
2010 J jnl
J. Comput. Sci. Coll.
Michael Sands, James Evans, Glenn David Blank
2009 J jnl
BMC Bioinform.
Merlin Veronika, James Evans, Paul Matsudaira, Roy E. Welsch, Jagath C. Rajapakse
2005 J jnl
Int. J. Inf. Manag.
James Evans, Laurence D. Brooks
2005 J jnl
Inf. Soc.
James Evans, Laurence D. Brooks
2003 conf
ISICT
Jacqueline Brodie, James Evans, Laurence D. Brooks, Mark J. Perry
2001 conf
VTC Fall
Jin Wang, Michael Caggiano, James Evans
1989 J jnl
Proc. IEEE
James Evans, Donald Turnbull
1986 J jnl
IEEE Trans. Syst. Man Cybern.
William Ernest Leigh, James Evans
1976 J jnl
Inf. Sci.
James Evans, Paul Kersten, Ludwik Kurz
redb/extractors/elf_extractors/elf_notes.py
← Index redb/extractors/elf_extractors/elf_notes.py python
import inspect
import binascii
from datetime import datetime, timezone
from typing import Any, List, Dict

from elftools.elf.elffile import ELFFile
from elftools.common.exceptions import ELFError

from redb.extractors.enum import Tag
from redb.extractors.elf_extractor import ELFExtractor
from redb.models.dataclasses import ELFNote


class ELFNotesExtractor(ELFExtractor):

    def __init__(
        self,
        filepath,
        log,
        exporters=None,
        index_prefix=None,
        elastic_index=None,
        known_benign=False,
        known_malicious=False,
        elf=None,
    ):
        super().__init__(
            filepath,
            log,
            exporters,
            index_prefix,
            elastic_index,
            known_benign,
            known_malicious,
            elf,
        )
        self.elf_notes = []
        self.elastic_index = self.index_prefix + "-elf_notes"
        self.log.debug(inspect.currentframe().f_code.co_name)

    def _get_note_type_string(self, note_type: int, note_name: str) -> str:
        """Convert note type number to human-readable string."""

        # GNU-specific note types
        if note_name == "GNU":
            gnu_types = {
                1: "NT_GNU_ABI_TAG",
                2: "NT_GNU_HWCAP",
                3: "NT_GNU_BUILD_ID",
                4: "NT_GNU_GOLD_VERSION",
                5: "NT_GNU_PROPERTY_TYPE_0"
            }
            return gnu_types.get(note_type, f"NT_GNU_UNKNOWN_{note_type}")

        # Generic note types
        generic_types = {
            1: "NT_PRSTATUS",
            2: "NT_FPREGSET",
            3: "NT_PRPSINFO",
            4: "NT_TASKSTRUCT",
            5: "NT_AUXV",
            6: "NT_PSTATUS",
            7: "NT_FPREGS",
            8: "NT_PSINFO",
            9: "NT_PRCRED",
            10: "NT_UTSNAME",
            11: "NT_LWPSTATUS",
            12: "NT_LWPSINFO",
            13: "NT_PRFPXREG"
        }

        return generic_types.get(note_type, f"NT_UNKNOWN_{note_type}")

    def _format_note_description(self, note_desc, note_type: int, note_name: str) -> str:
        """Format note description based on type for human readability."""
        try:
            if not note_desc:
                return ""

            # Handle build ID specifically (common case)
            if note_name == "GNU" and note_type == 3:  # NT_GNU_BUILD_ID
                if isinstance(note_desc, bytes):
                    return binascii.hexlify(note_desc).decode('ascii')
                return str(note_desc)

            # Handle ABI tag
            if note_name == "GNU" and note_type == 1:  # NT_GNU_ABI_TAG
                if isinstance(note_desc, bytes) and len(note_desc) >= 16:
                    # ABI tag contains OS, major, minor, subminor
                    import struct
                    try:
                        os_val, major, minor, subminor = struct.unpack('<IIII', note_desc[:16])
                        os_names = {0: "Linux", 1: "GNU", 2: "Solaris", 3: "FreeBSD"}
                        os_name = os_names.get(os_val, f"OS_{os_val}")
                        return f"{os_name} {major}.{minor}.{subminor}"
                    except:
                        pass

            # For binary data, convert to hex
            if isinstance(note_desc, bytes):
                # Limit size for very large descriptions
                if len(note_desc) > 256:
                    return binascii.hexlify(note_desc[:256]).decode('ascii') + "..."
                return binascii.hexlify(note_desc).decode('ascii')

            # For string data
            if isinstance(note_desc, str):
                return note_desc

            # Fallback
            return str(note_desc)

        except Exception as e:
            self.log.error(f"Error formatting note description: {e}")
            return str(note_desc) if note_desc else ""

    def _extract_note_data(self, note, section_name: str) -> ELFNote:
        """Extract data from a single note entry."""
        try:
            # Get note properties
            note_name = note.get('n_name', '').rstrip('\x00') if note.get('n_name') else ""
            note_type_raw = note.get('n_type', 0)
            note_desc_raw = note.get('n_desc', b'')

            # Handle note_type - pyelftools may return string or int
            if isinstance(note_type_raw, str):
                # pyelftools returned the type as a string like 'NT_GNU_BUILD_ID'
                note_type_str = note_type_raw
                # Map known string types to integers
                note_type_map = {
                    'NT_GNU_ABI_TAG': 1,
                    'NT_GNU_HWCAP': 2,
                    'NT_GNU_BUILD_ID': 3,
                    'NT_GNU_GOLD_VERSION': 4,
                    'NT_GNU_PROPERTY_TYPE_0': 5,
                    'NT_PRSTATUS': 1,
                    'NT_FPREGSET': 2,
                    'NT_PRPSINFO': 3,
                    'NT_TASKSTRUCT': 4,
                    'NT_AUXV': 5,
                    'NT_PSTATUS': 6,
                    'NT_FPREGS': 7,
                    'NT_PSINFO': 8,
                    'NT_PRCRED': 9,
                    'NT_UTSNAME': 10,
                    'NT_LWPSTATUS': 11,
                    'NT_LWPSINFO': 12,
                    'NT_PRFPXREG': 13,
                }
                note_type = note_type_map.get(note_type_raw, 0)
            else:
                note_type = note_type_raw
                # Get human-readable type string
                note_type_str = self._get_note_type_string(note_type, note_name)

            # Format description
            note_desc = self._format_note_description(note_desc_raw, note_type, note_name)

            return ELFNote(
                note_name=note_name,
                note_type=note_type,
                note_type_str=note_type_str,
                note_desc=note_desc,
                note_section=section_name
            )

        except Exception as e:
            self.log.error(f"Error extracting note data: {e}")
            return None

    def _extract_notes_from_sections(self, elf) -> List[Dict]:
        """Extract notes from note sections."""
        notes = []

        try:
            # Look for note sections
            for section in elf.iter_sections():
                if (section.name and
                    section.name.startswith('.note') and
                    hasattr(section, 'iter_notes')):

                    section_name = section.name
                    try:
                        for note in section.iter_notes():
                            note_data = self._extract_note_data(note, section_name)
                            if note_data:
                                notes.append(note_data)
                    except Exception as e:
                        self.log.debug(f"Could not process notes in section {section_name}: {e}")

        except Exception as e:
            self.log.error(f"Error extracting notes from sections: {e}")

        return notes

    def _extract_notes_from_segments(self, elf) -> List[Dict]:
        """Extract notes from PT_NOTE segments."""
        notes = []

        try:
            # Look for PT_NOTE segments
            for segment in elf.iter_segments():
                if segment.header.get('p_type') == 'PT_NOTE':
                    segment_name = f"PT_NOTE_segment_{segment.header.get('p_offset', 0)}"

                    try:
                        if hasattr(segment, 'iter_notes'):
                            for note in segment.iter_notes():
                                note_data = self._extract_note_data(note, segment_name)
                                if note_data:
                                    notes.append(note_data)
                    except Exception as e:
                        self.log.debug(f"Could not process notes in segment: {e}")

        except Exception as e:
            self.log.error(f"Error extracting notes from segments: {e}")

        return notes

    def tag(self):
        return Tag.ELF_NOTES.value if hasattr(Tag, 'ELF_NOTES') else "elf_notes"

    def extract(self):
        try:
            self.log.debug(inspect.currentframe().f_code.co_name)

            def extract_data(elf):
                all_notes = []

                # Extract notes from note sections
                section_notes = self._extract_notes_from_sections(elf)
                all_notes.extend(section_notes)

                # Extract notes from PT_NOTE segments
                segment_notes = self._extract_notes_from_segments(elf)
                all_notes.extend(segment_notes)

                # Remove duplicates (same note might appear in section and segment)
                unique_notes = []
                seen_notes = set()
                for note in all_notes:
                    note_key = (note.note_name, note.note_type, note.note_desc)
                    if note_key not in seen_notes:
                        seen_notes.add(note_key)
                        unique_notes.append(note)

                return unique_notes

            if not self._is_elf_file():
                return None

            result = self._with_elf_file(extract_data)
            if result is None:
                return None

            self.elf_notes = result
            return self.elf_notes

        except Exception as e:
            self.log.error(f"Error extracting ELF notes {self.hash.sha256}: {e}")
            return None

    def prepare_export_data(self, exporter_type: str) -> Any:
        self.log.debug(inspect.currentframe().f_code.co_name)

        if exporter_type == "ElasticsearchExporter":
            return self.elf_notes
        elif exporter_type == "ClickHouseExporter":
            try:
                # Return valid empty structure if no notes found
                # None is reserved for actual errors

                # Prepare data arrays for all notes
                data = []
                current_time = datetime.now(timezone.utc)
                for note in self.elf_notes:
                    row = [
                        self.sha256,
                        self.md5,
                        self.sha1,
                        note.note_name,
                        note.note_type,
                        note.note_type_str,
                        note.note_desc,
                        note.note_section,
                        current_time
                    ]
                    data.append(row)

                column_names = [
                    'sha256', 'md5', 'sha1',
                    'note_name', 'note_type', 'note_type_str',
                    'note_desc', 'note_section',
                    'analysis_date'
                ]

                if not data:
                    return None

                column_type_names = [
                    'FixedString(64)', 'FixedString(32)', 'FixedString(40)',
                    'LowCardinality(String)', 'UInt32', 'LowCardinality(String)',
                    'String CODEC(ZSTD(3))', 'LowCardinality(String)',
                    'DateTime64(3, \'UTC\')'
                ]

                return (data, column_names, column_type_names)

            except Exception as e:
                self.log.error(f"Error preparing export data: {e}")
                raise

    def get_clickhouse_table(self) -> str:
        return "redb_elf_notes"