Author Archives: admin

More XProc

I’ve been busy reading up on XProc today while walking through W3C’s XProc Test Suite.

An XML pipeline language has been on my wish list ever since my friend Henrik MÃ¥rtensson wrote something called eXtensible Filter Objects (XFO), an XML pipeline language not unlike XProc, about ten years ago and then lost interest, focussing instead on lean theories, business management and such. Some time before he moved on he wrote a Perl implementation of XFO and another friend, David Rosell, wrote a Java version of that, but unfortunate circumstances killed it all after XFO had been implemented for a few of our then-clients at Information & Media.

XProc, of course, does more than XFO ever did, but the ideas are the same. XProc is scratching a persistent itch for me and might (IMO, of course) very well become one of XML’s most important specs to date. For someone like me who is basically a non-programmer, being more of a markup theorist and dochead (to follow Ken Holman’s labelling of the degrees of XML geekery), it’s a wish come true.

Today, in spite of me going through the test suite and reading the spec, I feel that my most important action towards XProc wisdom was to check with Norman Walsh if he’s working on an XProc book yet (he is).

I’m getting there, though. I hope to finish a working pipeline for Cassis TI publishing tomorrow.

XProc

Leave a reply

I’m going to spend the next week or two doing a test implementation of XProc for our document management system, Cassis TI. XProc, as some of you will know, is a pipeline processing language for XML processing, in the same vein as pipe processing in the *nix world. It’s intended to standardise and ease XML processing by treating the processing as a black box consisting of smaller black boxes; in other words, what is inside is less interesting than how the in- and outputs are defined and used.

The test is about producing PDF output so it’s nothing fancy or new, but it’s important because I believe we can replace our current backend with an XProc-based processor, making things easier, faster and better for programmers and users alike.

Mobile Sync, Part Two

Leave a reply

I have an older IBM Thinkpad (a T42p) laptop with Ubuntu Studio installed. In version 9.10, syncevolution worked like a charm. All I had to do was to install, setup the N900 and sync, no problems whatsoever. Then I got brave and upgraded the laptop to Ubuntu 10.04 and syncevolution to the latest version.

Fail to sync.

And mind you, it doesn’t tell me what’s wrong, it just fails. I’ve tried installing older syncevolution packages, resetting bluetooth stuff, sacrificing my firstborn… nothing helps!

If you know what’s wrong, please let me know.

Roland D-50

Leave a reply

Got my hands on this venerable synth. Seems I’ll have to do repairs but oh, I’m so looking forward to this.

Mobile Sync

Leave a reply

After years of not being able to sync my Nokia mobile(s) with my Debian Linux desktop, syncevolution and the Evolution “groupware suite” have finally made that possible. I’ve had success with both my older Symbian 60-based phone, N95-2, and my (Maemo-based) N900.

See www.syncevolution.org for details on how to do this. My Debian Sid box required the apt sources from that site (it seems that Sid is lagging behind, at least for now; they’ve packaged the last beta but the site includes the released 1.0 version), but otherwise the install and sync both went without a hitch.

VirtualBox

2 Replies

I’ve switched from KVM to VirtualBox for my virtualisation needs. My Debian laptop is hosting and right now there is a Windows 7 guest. Apart from some slowness, especially with shared folders (on extfs3), the whole thing works like a charm. I can run XMetaL in the VirtualBox with no problems.

Finally it looks like I won’t be needing a Windows partition at work.

DITA Lists, Part Two

Leave a reply

Today it occurred me to have a look at the DITA Architecture Specification source to see how the people behind the spec would tag a list; as some of you know, this was the subject of my recent blog entry. There are a number of lists in that spec, many with introductory paragraphs, so it’s a pretty obvious way to find out, right? Well, after the examples in that spec, maybe.

Anyway, this is how they do it:

Introductory para:
<ul>
<li>Item</li>
<li>Item</li>
</ul>

This was one of my guesses, and I have to say that it’s better than any of the alternatives I could come up with. It’s not good markup, though, in my opinion, as it says that semantically, a paragraph is sort of a block-level superclass, a do-it-all and one that you must use if you need that introduction.

But then, why limit yourself to lists? Why aren’t notes tagged like that? Or definition lists, or images, or tables? Think about it. Doesn’t this feel just a little bit like a cop-out to you? It does to me. It feels like the author realised that he needed that wrapper but there was nothing he could cling to, other than this construction.

I’m not saying that my way is the only way (obviously it’s not) but this bothers me because it muddies the semantic waters.

BP Spilling Coffee

Leave a reply

http://mag.ma/andrew/617364

List Modelling

Leave a reply

I’ve been reading up on DITA. I’ve looked at the specs and the DTD before, obviously, but more from the perspective of an innocent bystander. The DTDs I implement in authoring systems and elsewhere are usually my own, and whenever I need to deliver content in some other format, I simply convert to it. This time things are a bit different, however, as we are considering doing a “DITA Edition” of the content management system I’m responsible for at work, and I need to know how DITA can fit into our stuff.

DITA’s got lots of things that I like, such as the combining of topic IDs with target IDs in references to avoid ID collisions. The DITA way is a very elegant solution and probably a better one than what I would usually do, which is to (in various ways in the DTD and in the authoring environment) make sure that authors can never end up in situations like it to begin with. There’s other stuff, too, but those are best left to another blog entry at some point.

Here, I want to talk about list modelling and specifically something that not only DITA but so many other DTDs and schemas seem to ignore, and that, in my mind, results in bad markup. Let’s start by discussing list semantics first:

A list is, well, a list of things. There are several types of lists, of which unordered and ordered are the most common, and the semantics are probably clear enough: the former lists stuff without a specific order (say, grocery lists) and the latter items whose order is significant (for example, David Letterman’s top ten lists). There’s also the definition list (which, in my mind, is not a list at all but a special case of a table, namely a two-column one), and probably some other types as well. In DITA, you can find something called “simple list”, which claims to limit what’s listed to one line per item, tops, without bullets or numbers, but to me that’s less about semantics and more about presentation.

So here’s a typical DITA list (HTML, DocBook and quite a few others look exactly like it, too):

<ul>
<li>Apples</li>
<li>Oranges</li>
<li>Bananas</li>
</ul>

There’s more to list semantics, though, at least in my mind. If you wanted to find a complete list in a document, you’d probably want to include its qualifying introduction (“Here’s the groceries you need to buy:”), and any and all information that goes between list items without being part of them but still belonging to the list as a whole. If your spouse is kind enough to subcategorise the grocery list to vegetables, fruit, dairy products and so on (I know I need the help), we’d have a multi-part list where the participating lists are part of a larger whole.

The introductory paragraph is where it gets tricky in DITA and similar structures. There are a LOT of block-level elements to choose from, but you cannot easily do a list that meets these requirements. This one, the preferred DITA way (at least if we choose to believe the examples in the spec), lacks a wrapper that identifies the list as one unit instead of a loose paragraph that happens to be followed by a list:

The fruit we need for tonight:
<ul>
<li>Apples</li>
<li>Oranges</li>
<li>Bananas</li>
</ul>
And the vegetables for tomorrow:
<ul>
<li>Cucumbers</li>
<li>Tomatoes</li>
</ul>

Of course, one could argue that our grocery list is really a section, but I would argue that the introductory paragraph is actually part of the list, but not necessarily a part of the whole section. What if I wanted to include images or perhaps a note to that section? Semantically, I can think of dozens of ways to reasonably expand the structure of such a surrounding section and still keep it on topic (that is, limiting it to subject matters concerning that central grocery list).

Keeping with DITA’s topic-based approach, we could certainly use a number of such sections and wrap the whole thing in a topic, but me, I think that’s overkill. All I want to do is include an introductory paragraph.

This, of course, is where some will argue that the introductory paragraph is really a heading. Definition lists in DITA and some other DTDs actually do have a heading for this very purpose, which to me hints that somebody did touch the subject at hand at some point, but then why do the “ordinary” lists without that heading? And of course, me, I think that introduction is not a heading at all, only a qualifier for the list.

Another option in DITA and others is to use the element as a wrapper:

The fruit we need for tonight:
<ul>
<li>Apples</li>
<li>Oranges</li>
<li>Bananas</li>
</ul>
And the vegetables for tomorrow:
<ul>
<li>Cucumbers</li>
<li>Tomatoes</li>
</ul>

This is perfectly valid, of course, but it ruins the intent of the element and creates a very odd (and ugly) mixed content that would be difficult to process properly.

What I would like to see is more in the lines of this:

<ul>
The fruit we need for tonight:
<li>Apples</li>
<li>Oranges</li>
<li>Bananas</li>
And the vegetables for tomorrow:
<li>Cucumbers</li>
<li>Tomatoes</li>
</ul>

Now we have a single list (our grocery list) that includes the necessary introduction(s). Of course, it’s still somewhat ugly; I, for one, dislike the relative lack of list item structure–I’d much rather see an item modelled more properly, perhaps divided into paragraphs and other block-level content, where the concepts block level and inline remain properly separated.

4G?

Leave a reply

Apple unveiled their newest iPhone, the 4G, earlier today. Looks like they are getting a little closer to what my Nokia N900 can do. I have to admit that they know how to market their stuff. If the functionality matched the hype, I might even be interested.

sgmlguru.org

on markup, film projection and more!

Author Archives: admin

More XProc

XProc

Mobile Sync, Part Two

Roland D-50

Mobile Sync

VirtualBox

DITA Lists, Part Two

BP Spilling Coffee

List Modelling

4G?