Showing posts with label Actor-Network-Theory. Show all posts
Showing posts with label Actor-Network-Theory. Show all posts

Friday, December 20, 2013

A Datawarehouse for Social Media for Learning


Recently i started to redesign our Mediabase, the database that includes data crawled from social media. A data warehouse model is more convenient for analyzing social media. Here i'm going to explain why.

All in all, the Mediabase design is not wrong. Behind it, its creators put the actor network theory (ANT) that claims all items in a system as actors. The following picture will help to understand how this idea applied to the Mediabase.


The Mediabase was an empty database with set of defined tables. Now it is filled on a regular basis with the help of watcher scripts. A watcher crawls one of Media: blogs, feeds, podctasts, e-mail archives. The traces left in a medium by its users consist of texts, tags or labels, urls and many other components. We collect these Artifacts in the Mediabase to be able later to apply them for different analytic tasks. Moreover, we examine users, e.g. e-mail senders, blog owners, forum posts authors and others. We can apply different analysis investigating changes of behaviors of uniquely defined users. Such an analysis shows activity patterns of users, their roles in communities, their interests and their closed peers in networks.

The Artifact and Medium are actors according to the ANT. The Community, the Agent (or User) and the Process are actors as well. All they have an influence on activities in online networks. That is why the Mediabase aim is to capture these changes.

But why do we need to collect all these data? Let us follow a teacher that aims to understand what problems do students have by solving an exercise. He can try to find an answer by observing online student communities and finding answers to the following questions:
  • Are there any questions about the exercise?
  • How many people have viewed the topic?
  • How many people have took part in the discussions?
  • What resources were used? How active are the students that talk about the exercise?
  • What are the sentiments and intents expressed in the discussions?

To find the answers can be time-consuming and impossible manually (depending on the number of students and the number of involved online communities). 1) Communities have unique structures: the positions of users in structures specify roles that represent community experts, communicative members and brokers which are members connecting several groups. 2) Huge amount of texts include relevant and irrelevant for learning information. It may include interesting explanations and links but can be overlooked by a person.

So the help of machines is needed not only because the communities are big. We are talking about situations when teachers, managers or stakeholders of online learning communities want to get answers to different questions about different time intervals of their communities.

Therefore, we need a database model that provides the right information in the right place at the right time with the right cost in order to support the right decision. In the traditional databases we get a data item by specifying two parameters, e.g. specifying the time and the medium parameters we get posts appeared in the medium in the given time period. Or we get users that were active at the given time period in the given medium. But more sophisticated queries require some time to execute because of joins, etc. Moreover, analyzing communities researchers have to mine data and afterwards make some experiments to find patterns.

To simplify the life cf researchers, I propose a data cube that is a multidimensional model with a set of measures - facts. We can navigate over the data cube by specifying dimensions that correlate with Actors of ANT model. Particularly, we can choose a User that posted at a Time in a Medium so that we get the output that characterize the User in the given Time point in the given Medium. But in this case the output includes different measures captured by mining social networks, texts or user activities.

On the following picture you can see the hierarchy of dimensions to the proposed data cube. The construction of the Mediabase cube is still not ready, as it should reflect all actors we find important in the online networks. Thus I still need to elaborate Process, Agent(types), Community(types), Artifacts dimensions.

Wednesday, July 25, 2012

some notes during reading "Networks, Crowds and Markets" by David Easley and Jon Kleinberg - Introduction

a network is a pattern of interconnections among a set of things

adding resources in a network can undermine its efficiency - Braess's paradox

a notion of equilibrium - a state that is self-reinforcing in that it provides no individual with an incentive to unilaterally change his or her strategy, even knowing how others will behave

we have a fundamental inclination to behave as we see other to behave: 1) the behavior of others convey the information or it is a direct benefit from aligning your behavior with that of others

in many cases you care more about aligning your own behavior with the behavior of your immediate neigbors  in the social network, rather than with the population as a whole.

a new behavior starts with a small set of initial adopters and then spread radially outward through the network.

a diffusion of technologies can be blocked by the boundary of a densely connected cluster in the network- a "closed community" of individuals who have a high amount of linkage among themselves, and hence are resistant to outside influences,



Tuesday, July 29, 2008

How can our experiences be followed?

Several days ago i read somewhere that experiences influence our emotions. That is a direct alarm for me to summarize that all operations we are performing in our life, at least virtual one, play roles in emotions we are producing. Therefore I can conclude that those two notions, operations (participatory design methods, media processes, user actions - call them as you want) and emotions are connected notions. One can either follow the actions accompanied with emotions like it was done in Situation Modeling and Smart Context Retrieval with Semantic Web Technology and Conflict Resolution or differentiate between those as different actors of a community system (when presented like an actor network model).

Some scholars follow the experiences through ontologies that help to summarize semantics in an ontology object and highlight connections between different entries. The approach applied in PALADIN: A Pattern Based Approach to Knowledge Discovery in Digital Social Networks uses a pattern approach. One can say that patterns can be like ontology objects when those are defined by the semantics and relationships between patterns are specified. Although a higher level of abstaction(for ontologies are concepts) is lacking, it is a moot point. The hierarchy can be specified by relations. Patterns exceed onlologies by a relatively easy language. So that it is natural for a user to construct a semantic knowledge base for his needs.

Wednesday, July 2, 2008

Media operations

Since i finished my master thesis a lot of questions left opened for me. One of those is the observation of media operations, somebody may call it events, others participatory design methods.



I defined the following table:



according to media centric theory of learning based on the
"The Knowledge-Creating company:How Japanese Companies Create the Dynamics of Innovation" from Nonaka and Takeuchi
and
"Communities of Practice: Learning, Meaning and Identity" from Wenger.

The theory was utilized in "Do you know a similar project I can learn from?" Self-monitoring of Communities of Practice in the Cultural Sciences by Klamma, Spaniol and Jarke

The engaged actors are those presented on the model:


Today I find very interesting paper in the same direction. The scholars tries to discover communities based on mutual awareness: Blog Community Discovery and Evolution Based on Mutual Awareness Expansion by Lin, Sundaram, Chi, Tatemura and Tseng. I found a minor number of resources to the topic thus each new paper is a treasury.