PubMed دسترسی آزاد

A divide and conquer strategy for recapitulating whole genome 3D structure using Hi-C data.

استودیوی صوتی مقاله

پخش حرفه‌ای فارسی و انگلیسی

در حال بررسی نسخه‌های صوتی ذخیره‌شده…

صوت تولیدشده با هوش مصنوعی است. برای کاربرد علمی یا درمانی، متن و منبع اصلی را بررسی کنید.
خواندن هوشمند فارسی و انگلیسی در حال آماده‌سازی صداهای مرورگر…
تنظیم صدای طبیعی و سرعت

صداهایی که در نامشان «Natural»، «Neural» یا «Online» دیده می‌شود معمولاً طبیعی‌ترند. انتخاب صدا به صداهای نصب‌شده در ویندوز و مرورگر شما بستگی دارد.

چکیده اصلی

The three dimensional (3D) spatial organization of the genome is closely linked to biological functions and can be captured by Hi-C assays through interrogating genome-wide chromatin interactions. Methodologies for inferring 3D structures from Hi-C data summarized as a two-dimensional (2D) contact matrix can be broadly placed within the paradigms of optimization-based and sampling-based. Many optimization-based methods are capable of constructing whole genome 3D structures but do not account for spatial dependency in the 2D data matrix nor cell heterogeneity in bulk Hi-C data, which provide an average over millions of cells. Sampling-based methods, on the other hand, are probabilistic model-based and can account for not only dependency, heterogeneity, but also other features inherent in Hi-C data, such as over-dispersion and sparsity. However, whole-genome 3D structure recapitulation is too computationally expensive for sampling-based methods, while chromosome-by-chromosome strategies for sampling-based methods ignore important information on inter-chromosomal contacts. To address these issues, we propose the truncated Random effect EXpression-cut and paste (tREX-cap) method, which applies the tREX model within a divide and conquer strategy. The resulting method inherits the good data-feature-cognizant properties of tREX and, in the meantime, can efficiently infer the whole genome 3D structure. We demonstrate the performance of tREX-cap through an extensive simulation study and analyses of a Hi-C lymphoblastoid dataset and a Hi-C IMR90 dataset.

متن کامل اصلی

نسخه دارای مجوز در منبع علمی در دسترس است.

لینک مستقیم از metadata منبع گرفته شده و در تب جدید باز می‌شود.

باز کردن متن کامل

کلیدواژه‌ها

over-dispersionsparsityspatial dependency
در همین زیرشاخه

مقاله‌های مرتبط

PubMed2026

Preparedness of the Ghana Health Service for field epidemiology and applied biostatistics: a systematic review protocol of infectious disease surveillance, outbreak investigation methodologies, and statistical modeling capacities in resource-limited settings.

BACKGROUND: Infectious disease outbreaks pose significant threats to global health security, with resource-limited settings in West Africa bearing a disproportionate burden. Despite sustained investments in field epidemiology training and surveillance system strengthening, no comprehensive systematic synthesis exists of Ghana Health Service preparedness for field epidemiology and applied biostatistics. This protocol addresses the prima…

PubMed2026

Data-adaptive identification of effect modifiers through stochastic shift interventions and cross-validated targeted learning.

In epidemiology, identifying subpopulations that are particularly vulnerable to exposures and those who may benefit differently from exposure-reducing interventions is essential. Factors such as age, gender-specific vulnerabilities, and physiological states such as pregnancy are critical for policymakers when setting regulatory guidelines. However, current semiparametric methods for estimating heterogeneous treatment effects are often …

PubMed2026

Risk estimation and dynamic prediction using discrete-time joint models for longitudinal and multistate data with interval and state censoring.

This paper presents a joint model of multivariate longitudinal data and multistate data with application to modeling and predicting autoantibody development in The Environmental Determinants of Diabetes in the Young (TEDDY) study. The model quantifies the risks of state transitions based on observed time-varying and non-time-varying risk factors. Based on the estimated model, a dynamic prediction approach is suggested to predict future…