Mapping and Describing Geospatial Data to Generalize Complex Mapping and
Describing Geospatial Data to Generalize Complex Models: The Case of
LittoSIM-GEN Models
- URL: http://arxiv.org/abs/2101.07523v1
- Date: Tue, 19 Jan 2021 09:16:05 GMT
- Title: Mapping and Describing Geospatial Data to Generalize Complex Mapping and
Describing Geospatial Data to Generalize Complex Models: The Case of
LittoSIM-GEN Models
- Authors: Ahmed Laatabi, Nicolas Becu (LIENSs), Nicolas Marilleau (UMMISCO),
C\'ecilia Pignon-Mussaud (LIENSs), Marion Amalric (CITERES), X. Bertin
(LIENSs), Brice Anselme (PRODIG), Elise Beck (PACTE)
- Abstract summary: We provide a mapping approach to structure, describe, and automatize the integration of geospatial data into agent-based models.
This paper was part of the LittoSIM-GEN project.
- Score: 0.0
- License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/
- Abstract: For some scientific questions, empirical data are essential to develop
reliable simulation models. These data usually come from different sources with
diverse and heterogeneous formats. The design of complex data-driven models is
often shaped by the structure of the data available in research projects.
Hence, applying such models to other case studies requires either to get
similar data or to transform new data to fit the model inputs. It is the case
of agent-based models (ABMs) that use advanced data structures such as
Geographic Information Systems data. We faced this problem in the LittoSIM-GEN
project when generalizing our participatory flooding model (LittoSIM) to new
territories. From this experience, we provide a mapping approach to structure,
describe, and automatize the integration of geospatial data into ABMs.
Related papers
- Improving Full Waveform Inversion in Large Model Era [25.26004497243484]
We show that a model trained entirely on simulated and relatively simple data can generalize remarkably well to challenging geological benchmarks.<n>Our model achieves state-of-the-art performance on OpenFWI and significantly narrows the generalization gap in data-driven FWI.
arXiv Detail & Related papers (2026-02-27T23:33:06Z) - A Highly Configurable Framework for Large-Scale Thermal Building Data Generation to drive Machine Learning Research [22.54521342959957]
BuilDa is designed to produce synthetic data of adequate quality and quantity for machine learning (ML) research.<n>It does not require profound building simulation knowledge to generate large volumes of data.<n>We demonstrate BuilDa by generating data and utilizing it for a transfer learning study involving the fine-tuning of 486 data-driven models.
arXiv Detail & Related papers (2025-11-29T13:31:02Z) - GEO-Bench-2: From Performance to Capability, Rethinking Evaluation in Geospatial AI [52.13138825802668]
GeoFMs are transforming Earth Observation, but evaluation lacks standardized protocols.<n> GEO-Bench-2 addresses this with a comprehensive framework spanning classification, segmentation, regression, object detection, and instance segmentation.<n>Code, data, and leaderboard for GEO-Bench-2 are publicly released under a permissive license.
arXiv Detail & Related papers (2025-11-19T17:45:02Z) - Generative Models for Synthetic Data: Transforming Data Mining in the GenAI Era [49.46005489386284]
This tutorial introduces the foundations and latest advances in synthetic data generation.<n> Attendees will gain actionable insights into leveraging generative synthetic data to enhance data mining research and practice.
arXiv Detail & Related papers (2025-08-27T05:04:07Z) - A Survey on Tabular Data Generation: Utility, Alignment, Fidelity, Privacy, and Beyond [53.56796220109518]
Different use cases demand synthetic data to comply with different requirements to be useful in practice.
Four types of requirements are reviewed: utility of the synthetic data, alignment of the synthetic data with domain-specific knowledge, statistical fidelity of the synthetic data distribution compared to the real data distribution, and privacy-preserving capabilities.
We discuss future directions for the field, along with opportunities to improve the current evaluation methods.
arXiv Detail & Related papers (2025-03-07T21:47:11Z) - Exploring the Landscape for Generative Sequence Models for Specialized Data Synthesis [0.0]
This paper introduces a novel approach that leverages three generative models of varying complexity to synthesize Malicious Network Traffic.
Our approach transforms numerical data into text, re-framing data generation as a language modeling task.
Our method surpasses state-of-the-art generative models in producing high-fidelity synthetic data.
arXiv Detail & Related papers (2024-11-04T09:51:10Z) - (Deep) Generative Geodesics [57.635187092922976]
We introduce a newian metric to assess the similarity between any two data points.
Our metric leads to the conceptual definition of generative distances and generative geodesics.
Their approximations are proven to converge to their true values under mild conditions.
arXiv Detail & Related papers (2024-07-15T21:14:02Z) - A Review of Modern Recommender Systems Using Generative Models (Gen-RecSys) [57.30228361181045]
This survey connects key advancements in recommender systems using Generative Models (Gen-RecSys)
It covers: interaction-driven generative models; the use of large language models (LLM) and textual data for natural language recommendation; and the integration of multimodal models for generating and processing images/videos in RS.
Our work highlights necessary paradigms for evaluating the impact and harm of Gen-RecSys and identifies open challenges.
arXiv Detail & Related papers (2024-03-31T06:57:57Z) - T1: Scaling Diffusion Probabilistic Fields to High-Resolution on Unified
Visual Modalities [69.16656086708291]
Diffusion Probabilistic Field (DPF) models the distribution of continuous functions defined over metric spaces.
We propose a new model comprising of a view-wise sampling algorithm to focus on local structure learning.
The model can be scaled to generate high-resolution data while unifying multiple modalities.
arXiv Detail & Related papers (2023-05-24T03:32:03Z) - Probabilistic Point Cloud Modeling via Self-Organizing Gaussian Mixture
Models [19.10047652180224]
We present a continuous probabilistic modeling methodology for spatial point cloud data using finite Gaussian Mixture Models (GMMs)
We use a self-organizing principle from information-theoretic learning to automatically adapt the complexity of the GMM model based on the relevant information in the sensor data.
The approach is evaluated against existing point cloud modeling techniques on real-world data with varying degrees of scene complexity.
arXiv Detail & Related papers (2023-01-31T19:28:00Z) - Towards a mathematical understanding of learning from few examples with
nonlinear feature maps [68.8204255655161]
We consider the problem of data classification where the training set consists of just a few data points.
We reveal key relationships between the geometry of an AI model's feature space, the structure of the underlying data distributions, and the model's generalisation capabilities.
arXiv Detail & Related papers (2022-11-07T14:52:58Z) - Modular machine learning-based elastoplasticity: generalization in the
context of limited data [0.0]
We discuss a hybrid framework that can work on a variable amount of data by relying on the modularity of the elastoplasticity formulation.
The discovered material models are found to not only interpolate well but also allow for accurate extrapolation in a thermodynamically consistent manner far outside the domain of the training data.
arXiv Detail & Related papers (2022-10-15T17:35:23Z) - Towards Understanding and Mitigating Dimensional Collapse in Heterogeneous Federated Learning [112.69497636932955]
Federated learning aims to train models across different clients without the sharing of data for privacy considerations.
We study how data heterogeneity affects the representations of the globally aggregated models.
We propose sc FedDecorr, a novel method that can effectively mitigate dimensional collapse in federated learning.
arXiv Detail & Related papers (2022-10-01T09:04:17Z) - Joint Gaussian Graphical Model Estimation: A Survey [31.811209829224293]
Graphs from complex systems often share a partial underlying structure across domains while retaining individual features.
Growing evidence shows that the shared structure across domains boosts the estimation power of graphs.
This manuscript surveys recent work on statistical inference of joint Gaussian graphical models.
arXiv Detail & Related papers (2021-10-19T21:56:27Z) - An Ample Approach to Data and Modeling [1.0152838128195467]
We describe a framework for modeling how models can be built that integrates concepts and methods from a wide range of fields.
The reference M* meta model framework is presented, which relies critically in associating whole datasets and respective models in terms of a strict equivalence relation.
Several considerations about how the developed framework can provide insights about data clustering, complexity, collaborative research, deep learning, and creativity are then presented.
arXiv Detail & Related papers (2021-10-05T01:26:09Z) - Dataset Cartography: Mapping and Diagnosing Datasets with Training
Dynamics [118.75207687144817]
We introduce Data Maps, a model-based tool to characterize and diagnose datasets.
We leverage a largely ignored source of information: the behavior of the model on individual instances during training.
Our results indicate that a shift in focus from quantity to quality of data could lead to robust models and improved out-of-distribution generalization.
arXiv Detail & Related papers (2020-09-22T20:19:41Z)
This list is automatically generated from the titles and abstracts of the papers in this site.
This site does not guarantee the quality of this site (including all information) and is not responsible for any consequences.